Even if really great, it still can’t use tools. The only consumer Google thing with access to MCP tools is Gemini Spark and that has lots of other problems. I wish they would combine their efforts on a great consumer product but it’s Google we’re talking about…
I still dream of the day that we get full tool parity in voice and text mode so your voice assistant can do everything you connect for you. Grok and Claude are btw almost there, only a very few minor built-in tools don’t exist in voice mode, but I already use both to connect to heaps of things! It’s so valuable to verbally discuss something with an agent, have the agent pull in context from GitHub, Notion, email, and then create artifacts somewhere
Note sure what you mean by "tools". It does support function calling, so you can hook it to whatever tools you want. I'm using it with Home Assistant to control my smart home devices.
> 1. Allow us to edit the photo order of the Shared Album (instead of the order that the photos were added to it). Currently, my wedding unfolds in the wrong order, and my kid turns 6 months old before being born.
Comet IS a normal chromium browser. The assistant thing is just a chat on the right sidebar and you can tell it stuff like “click around on this website to do X”, which it can then do for a limited amount of turns
I have a pass with Gemini 3.8 low over every PR Claude makes that specifically flags this. It points out all the slop comments, docs, commit messages. Doing this has greatly improved my comment and commit text quality
Funny thing is, Claude often “disagreed with part of the review and decided to not adopt the requested changes” lol
More because Spark is not good. Gemini Flash 3.8 is an awesome model! I use it all the time
Spark asks you for every single tool call, often forgets earlier parts of the conversation, is unable to use apps unless you explicitly tag them and and and
Google has great models, but their harnesses and products are just bad
Google definitely does not have good models and if it does, Gemini Flash series models are definitely not them. I use the Google Gemini app daily (for searches, research etc.), everyday, for some query, gemini flash 3.8 hallucinates, does not understand context etc. in quite short conversations. I switch back to 3.1 pro and it works fine.
Which brings me to why do they get good benchmark scores? Have never really understood this. They must be optimizing just for this.
I used IntelliScript a lot. I give Claude a subtitle export, ask it to cut it and import the cut subtitle track with IntelliScript. Resolve then creates a video version that’s cut based on the subtitle changes
I’m also doing YouTube and the first thing I did was create a separate Google account for it. I’d be very surprised if Theo has everything running on the same personal account
All of the subscription AI platforms are trimming down quotas across the board to push users into higher tiers. Whatever they can do. Local inference needs to meet pricing sooner
Personally jumping around a lot to get a feeling for exactly that, and these days liking the Grok harness out of all of them the most
reply