Typesafe AI raises $870M at $7.5B
https://typesafe.ai/blog/series-aiEDIT: As I wrote this Microsoft just released their own Decision-1 model [3].
[1] https://developers.openai.com/api/docs/guides/decisions
[2] https://unsloth.ai/docs/basics/train-your-own-decision-model...
[3] https://commandline.microsoft.com/microsoft-decision-1-model...
They may well be a good team to throw money behind if you are hoping to bet on a new AI lab.
I see a lot of people parroting the quick open source alternatives as being better on the benchmarks, but it's such a new category that I'm not convinced we have solid benchmarks.
I'm hoping a company releases an internal eval benchmark for these options. I'm sure some of the open source ones are solid in some cases, but would love to see more reliable data.
Invariably near-AGI systems created by OpenAI/Anthropic will be very destabalizing. In the end the world will probably regulate AI capable of [any] <-> [any] input/output types. Models will need to be limited on their outputs by law so they cannot have unbounded, unpredictable outcomes. Jev is the ideal version of "benefits of AI without making humans obsolete" that might be the consensus once the track superhuman AI and its consequences are clear.
If being an “AI Researcher” is a ticket to multimillion dollar salary, AI training talent cannot be contained to a handful of companies. It’ll become more common and diffuse. The old advice of not fine tuning, because it’s hard, goes out the window as that knowledge diffuses through the industry.
A similar thing is happening in search. For a long time labs have trained tailored embedding models. And now companies like SID training their own agentic models that are smaller and faster at search than GPT-5.
Unless china takes leadership in frontier space the picture is next :
1. cheap workhorses for classification, routing, other scenarios : Jev 2. coding agents with less erros : Anthropic/Openai, etc. 3. Science /Legal/Medical : A mixture of Jev+Anthropic scenarios
You won the competition with VCs
We released an Apache-2.0, open-weight 4B decision model that scores above Jev 1.13 on JevBench's composite score (72.5 vs 71.5) and is currently the top open model there: https://benchmarkheaven.com/jev-models . Newer models coming even larger than beat Jev in intelligence as well.
- Same contract as Jev: state + typed questions in, calibrated probabilities out, one forward pass, no generated tokens. - Your data never leaves your environment, and there's no per-call fee.
Weights, card and run instructions: https://huggingface.co/h2oai/h2o-lightning-4b