Skip to main content

4 posts tagged with "models"

View all tags

Jev: a decision model you put in front of your LLMs to route traffic

· 7 min read
Rafael Fernandes
NLP Engineer & Tech Writer at WiLine
Share:
Routing · AI News

Jev decides, your LLMs answer

A request comes inThe right model, in 127 ms

TypeSafe shipped its first model on 15 September, and the interesting thing about Jev is what it refuses to do. It doesn't write you a paragraph. You give it a request and it hands back one structured value — a label, a class, a decision — in, they say, 70 to 500 ms. Founder Diogo Almeida's framing is the clearest line in the post: "Think of Jev as a frontier-intelligence function call: unstructured state in, typed probabilistic decisions out."

Most models are built to talk to people. Jev is built to be called by code — and the first job that shape fits is routing.

Muse Glimmer: a 30B agentic model that runs on one GPU, no data center required

· 4 min read
Rafael Fernandes
NLP Engineer & Tech Writer at WiLine
Share:
Models · AI News

A serious agent, no data center required

Cloud-only agentsOne GPU, fully local

Most "run it locally" model announcements come with an asterisk — smaller, weaker, a toy version of the real thing. Meta's newest release doesn't: Muse Glimmer, a 30B multimodal model built specifically for agentic work, fits on a single consumer GPU and beats larger models on the benchmarks that actually measure agent behavior.

GPT-5.6: OpenAI's new pitch is cheaper per task, not just smarter — verify it on your workload before you switch

· 10 min read
Rafael Fernandes
NLP Engineer & Tech Writer at WiLine
Share:
Models · AI News

The frontier race just changed lanes: from smarter to cheaper per task

Benchmark pointsToken economics

OpenAI shipped GPT-5.6 on July 9 — a family of three models (Luna, Terra, Sol) — and the headline claim isn't a leaderboard score. It's an efficiency number: frontier coding performance on less than half the output tokens. If you build agents, that's a claim about your bill, not about bragging rights. It's also exactly the kind of claim you should measure yourself.

GLM-5.2: the only open-weight model in the top 10 — and you can run it on WEC

· 4 min read
Rafael Fernandes
NLP Engineer & Tech Writer at WiLine
Share:
Models · AI News

An open-weight model just cracked the proprietary top 10

Closed frontierOpen weights

Look at almost any current model leaderboard and the top is a wall of Anthropic and OpenAI. Then, sitting in the top 10, there's one outlier that isn't proprietary at all: GLM-5.2 from Z.ai — open weights, MIT-licensed. That's the story worth paying attention to.