Probably Raises $9M to Build AI Systems That Catch Their Own Errors
Andreessen Horowitz-backed startup Probably uses validator systems and smaller models to achieve 99.99% accuracy in AI outputs, reducing hallucination risk and token costs.
Andreessen Horowitz-backed startup Probably uses validator systems and smaller models to achieve 99.99% accuracy in AI outputs, reducing hallucination risk and token costs.
Professional services firm pulls October 2025 report on agentic AI after organizations dispute accuracy of AI-usage claims.
OpenAI's new default ChatGPT model reportedly achieves a 52.5% reduction in hallucinated claims on high-stakes queries, grounded in real user-flagged failure data.
ChatGPT's new default model cuts fabricated claims by more than half on high-stakes prompts and shows users exactly what personal context shaped each response.