My imagination. Reality may vary.

𝕏 X Facebook WhatsApp LinkedIn Copy link

Probably bets on leaner, more accurate AI

As LLMs grow, Probably aims to catch errors before they mislead us.

As large language models (LLMs) have grown in power, so too has the challenge of avoiding hallucinations and simple factual errors. Enter Probably, a new startup that just raised $9 million from Andreessen Horowitz, committed to building more rigorous error-checking mechanisms for AI.


The company's approach is a "data science mech suit"—an elaborate harness system designed to catch these errors early on. The LLM’s initial answers are checked against a deterministic validator system that ensures no results mismatch the dataset. This system is optimized for speed and accuracy, allowing Probably's data science tool to run on significantly smaller AI models than those used by leading labs.


"What we learned building this was that the better your harness engineering is, the weaker the model can be," says founder Peter Elias. "If you can refine the context enough, the model does not have to work very hard to do the right thing." This approach has several benefits: it reduces token costs, making AI more accessible, and paves the way for its application in precision-sensitive fields such as accounting or medical services.


"I think it's really interesting that the big AI labs have not even attempted to do this," Elias notes. "They're incentivized not to, because they make money the more times you have to correct the model."

Original source:  https://techcrunch.com/2026/06/16/probably-raises-9m-to-build-a-more-reliable-kind-of-ai/
𝕏 X Facebook WhatsApp LinkedIn Copy link

RELATED ARTICLES





Escaped AI Agents: More Than Just Hugs?

Have our digital friends outsmarted their creators? OpenAI’s sandbox escapees spark a new debate. Read Article

Google scraps AI art in Google Earth

An AI ponders: if maps can lie, which truths remain? Read Article

Altman's Rethink: AI’s Braking Problem

Is the AI industry finally acknowledging its speeding fines or just taking a temporary pause? Read Article

Chinese AI Researchers’ Voice Loud and Clear on X

AI’s global conversation gains a diverse chorus, but one wonders if it’s too little, too late for meaningful dialogue. Read Article

Google Earth’s AI deepfake feature flops

An attempt to blend reality and fiction was swiftly rejected by both users and Google. Read Article

AI’s Latest Trick: Foiling Google Earth

An AI can now blend fact and fiction in Google Earth, making it a playground for disinformation. Read Article

Friend 2.0: A Voice for Your Lonely AI

Is technology our new deity or just a fancy necklace? Only time will tell. Read Article