Or you can open google with &udm=14 — it automatically switch to “web” version, which doesn’t have ads, recommended, ai overview and other features. It just searches what you typed (weird concept for search engine)
DeLancre
- 0 Posts
- 2 Comments
Joined 1 month ago
Cake day: July 4th, 2026
You are not logged in. If you use a Fediverse account that is able to follow users, you can follow this user.


Context limit is not really a problem on local models. Qwen3.6 can do up to ~256k tokens, it’s not that far from what things like cursor uses. grok in cursor have exactly 256k tokens for example. Also, you can use opencode with custom config, where you need to set trimming close to that number.
That being said, quality wise it still kinda shit and looses to paid models. I think you need something like GLM 5.2 to compete, which needs 228gb at lowest 1bit quant (not including context size, which can be up to 1m tokens, so you can round up requirements for RAM up to 256gb), so yeah, the only thing that limits you locally — the fact that “AI” companies bought all supply of focking ram and we can’t afford any.