Few technologies have climbed as fast as large language models, the systems behind ChatGPT, Claude, Gemini and their rivals.
Iternal Technologies (iternal.ai), the Austin-based enterprise AI strategy company, today announced Ultrabench (ultrabench.ai), a free AI benchmark aggregator that ranks every current large language ...
Large language models (LLMs) show promise in assisting knowledge-intensive fields such as oncology, where up-to-date information and multidisciplinary expertise are critical. Traditional LLMs risk ...
Loop scaling laws from Meta AI researchers predict that a Looped Mixture-of-Experts model can match a conventional MoE model ...
Modern factories are drowning in data but starving for knowledge. Sensors, maintenance logs, manuals, and fault reports pile ...
Salesforce Koa, the company's first CRM reasoning model built on NVIDIA Nemotron, launched at Dreamforce 2026 with self-administered benchmark claims -- while a Bloomberg investigation documents that ...
GPT-6 Astra launched September 3 with a 99.9% ARC-AGI-3 score and a 1.05 million-token context window, the largest ever on a ...
Large language models seem to be a double-edged sword. While they can answer questions -- including questions on how to create code and test it -- the answers to those questions are not always ...
Kimi K2.7 Code delivers a 21.8% improvement in real-world coding benchmarks, costing 13¢–78¢ per prompt with mixed speed and efficiency results. Moonshot AI’s baseline coding model set the foundation ...