There are new 3.6 and 3.5 models today, but Google is already training Gemini 4.
Your local LLM can use some help.
NVIDIA's PersonaPlex is fast, local, deeply impressive, and you can run it on just 8GB of VRAM.