LLM Responses compared
I thought a nice exercise would be to take a relatively simple prompt and assess how the Closed and Open models currently available compare. This is the prompt that I used:
<PROMPT>
I’m taking my daughter to an oral surgeon today to discuss removing her 3 present wisdom teeth. In overwhelming (?) circumstances wisdom teeth are removed as there is not room in the mouth for them. Is there ever going to be a chance mother nature starts creating humans without any wisdom teeth?
</PROMPT>
First, all the responses essentially touched on the same items. Some had more detail and more flourish in their delivery. Consider how I’m using this data; it has informed me and enlightened me on the question at hand. I have no intention of repurposing this copy on any site except for illustrative purposes, which is Chrisspeak for not caring as much about how “pretty” the prose of the LLMs are. Here’s a slideshow of the responses.
Here are a couple opinions in no particular order:
1. the Meta.ai response is sad, and they are the only Open source US model in the list.
2. Deepseek R1 and Qwen both have impressive results.
3. I like the Grok response more that I anticipated. But then again, I was predisposed not to like it (shakes fist at Elon) but I keep an open mind.
Musings – Prompting, productivity, and context
Prompting, Productivity and Context Finish the following sentence: "Blogging is so ..." and yet here I am. Prompting I've been trying to engage people close to me as to their AI experiences and uses, either professionally or personnally. I find myself reminding...
TEDs v 2.what?
I came across this TechCrunch article covering the recent news that Google is committing 150M to develop AI glasses with Warby Parker: https://techcrunch.com/2025/05/20/google-commits-150m-to-develop-ai-glasses-with-warby-parker/ My immediate thought was a new take on...
Bill Gates, not a jerk billionaire
Bill Gates has announced he's giving away 200B over the next 20 years to help address 3 moonshot global needs. Here's the announcement from Gates Foundation There are a lot of people that idolize billionaires and I am not one of them. A have an internal screed about...
Gemini = Lazy Google?
Now my mind is just equating Gemini to Lazy Google. For example, I wanted to modify the footer of tquist.com website. But what footer and where do I modify this? In the pre-Gemini (or LLM) days I would have typed my question into Google search and looked for...
AI Musings 5-5-2025
I am struggling to determine the best way to compare the big three LLMs. Side by side comparison of the same prompt is logical but yeesh, not sure I’ll have the time for such detailed analysis. Another thought I’ve had is to use a different one for time periods and...
AI Musings 5-2-2025
I have been trying to use the top 3 (in my mind at least) LLMs regularly. I feel most likely to grab Copilot as Microsoft has so nicely put a cool looking rainbow colored icon on the taskbar of Windows. I scratch my head a little as I don't recall allowing Microsoft...
AI limitations – let’s change a Google Voice number
TQuist had a Google Voice number for a long time. I never used it well, and then I neglected it, and then it was releasted back into the wild. I confirmed earlier my previous number has been assigned to some else, which makes sense due to my inactivity. So I want a...