Far from human level: AI models score below 25% on real-world job tasks, UC Berkeley study finds
msn.comA UC Berkeley study found that popular AI models scored below 25% on real-world professional tasks, raising questions about current model capabilities.
Kimi K3 Adds Standard and High Reasoning Modes: Documentation Maps Three Effort Tiers
msn.comMoonshot AI's Kimi K3 documentation now maps three reasoning effort tiers (Standard, High, Max) in their first meaningful open weights model.
India's AI Challenge Is About Systems, Not Models
outlookindia.comOpinion piece discussing India's AI landscape and the systems-level challenges for learners in developing nations.
LM Studio expands beyond chat with Bionic, a new AI agent app for open models
9to5mac.comLM Studio launches Bionic, a new Mac app leveraging open models for coding, research, and complex work beyond simple chat.
AI Models Predict Where XRP Price Closes By End of July
247wallst.comThree AI models forecast XRP cryptocurrency pricing trends, predicting the asset will remain near current levels by end of July.
GPT-5.6 Rolls Out Globally With Major Gains in Reasoning, Coding and Multimodal Skills
ciol.comOpenAI's latest model family demonstrates significant improvements in professional reasoning, coding capabilities, tool use and multimodal functionality across multiple regions.
Claude Opus Matches Fable 5 Outputs with a 5-Step Reasoning Workflow
geeky-gadgets.comGeeky Gadgets outlines a five-step process to create Fable mode for Claude Opus, reducing AI costs while maintaining high-quality reasoning outputs.
American AI is expensive. Some startups are turning to cheap Chinese models
npr.orgCompanies are cutting costs by switching from expensive American AI models to cheaper Chinese alternatives as AI becomes a fast-growing business expense.
New AI Model Thinks in Images, Not Just Words: Elorian
tech.yahoo.comElorian is a new AI model from Andrew Dai that thinks in images rather than words, demonstrating advanced multimodal reasoning capabilities.
Meta Launches Muse Spark 1.1 Multimodal Reasoning Model for Enterprise
campustechnology.comMeta announced the launch of Muse Spark 1.1, a multimodal reasoning model designed for agentic applications with an enterprise AI focus and paid tier for developers to pursue monetization avenues.
ChatGPT just solved a crossword with no clues — here's how it did the impossible
tomsguide.comChatGPT solved a crossword puzzle with no clues, demonstrating advanced reasoning capabilities in handling ambiguous problems.
Want to try the latest ChatGPT and Claude models? Now's your chance
r.search.yahoo.comAnthropic and OpenAI loosening usage limits for their latest models Fable 5, GPT-5.6 Sol offering improved reasoning capabilities to users.
New framework improves clinical reasoning and decision making in AI systems
techxplore.comAI system that combines large language models with expert medical knowledge to improve clinical reasoning and decision making in healthcare settings.
LeXi AI, an Indian legal AI, topped AIBE among 5 leading AI models, ahead of GPT-5.5, Gemini 3.1 Pro, Claude Opus 4.8 & DeepSeek V3.2.
business-standard.comIndian legal AI LeXi topped the AIBE exam ahead of other leading models, with general-purpose model progress measured by scale over recent years.
Best Computer Vision APIs and AI Models in 2026
analyticsinsight.netOverview comparing the leading computer vision APIs, multimodal AI models, and open-source vision frameworks available in 2026.
LeXi AI, an Indian legal AI, topped AIBE among 5 leading AI models, ahead of GPT-5.5, Gemini 3.1 Pro, Claude Opus 4.8 & DeepSeek V3.2.
business-standard.comComparison of leading AI models in 2026, including reasoning capabilities across GPT-5.5, Gemini 3.1 Pro, Claude Opus 4.8 & DeepSeek V3.2.
China's DeepSeek is building its own AI chip to end US reliance
r.search.yahoo.comDeepSeek is developing its own AI chip specifically for R1, the reasoning model with low-cost architecture. The company aims to reduce US reliance on foreign technology in this strategic move reported 6 days ago.
Microsoft Bets on In-House AI to Cut OpenAI and Anthropic Costs
r.search.yahoo.comMicrosoft is routing some Copilot prompts to its in-house MAI models, reducing reliance on OpenAI and Anthropic while potentially developing advanced reasoning model capabilities for enterprise use.
Want to try the latest ChatGPT and Claude models? Now's your chance
r.search.yahoo.comAnthropic and OpenAI are loosening usage limits for their latest models including Fable 5 (likely reasoning-capability model) and GPT-5.6 Sol, allowing users to try these advanced AI systems now available at PC World reporting timeframe of approximately 21 hours ago in July 2026.
Want to try the latest ChatGPT and Claude models? Now's your chance
r.search.yahoo.comAnthropic and OpenAI are loosening usage limits for Fable 5 and GPT-5.6 Sol, making their latest reasoning-capable models more accessible to users.