Search results for Claude
AI & Enterprise
Open-source benchmark Livenerf tracks whether AI model performance changes after release
An open-source benchmark called Livenerf has emerged to track over time whether an AI model’s performance declines after release. Built to detect changes caused by shifts in inference settings or routing, it aims to verify differences through repeated measurements rather than user perception. The current target is Claude Opus 5.5, with 78 questions selected for ongoing monitoring. The test does not represent overall API performance, and its developer says current data are insufficient to judge degradation.
AI & Enterprise
On-device AI is ready to shake up an AI market dominated by frontier models
AI expert Sigil Wen (시질 웬), who developed the on-device AI model Underdog, said AI models are moving out of data centres and into personal devices. He said most AI tokens will soon be generated on smartphones and laptops. He cited benchmark results showing a 27-billion-parameter Qwen 3.8 27B model on Macs and iPhones scoring higher than Opus 4.6 Max. Wen said large models are unnecessary for daily tasks and stressed speed and security as key advantages of on-device AI.
AI & Enterprise
Q3 semiconductors slide, software stocks rise; which stock Jim Cramer says can climb more
U.S. corporate software stocks posted the most notable rebound in the third quarter, CNBC reported, citing Jim Cramer’s review of the market. He said software shares returned to leadership as selling tied to artificial intelligence worries faded. The iShares Expanded Tech-Software Sector ETF rose 17 percent in the quarter, while the iShares Semiconductor ETF fell 11 percent. Cramer highlighted Salesforce and Microsoft and flagged rate hikes as his top concern for the new quarter.
-
AI & Enterprise
AI researchers warn of \'intelligence explosion\' as \'Mythos-level\' moments could come monthly
-
AI & Enterprise
Anthropic\'s Claude moves into U.S. federal agencies with government-only AI launch
-
AI & Enterprise
MSIT minister steps in amid questions over Korean AI foundation model project, two-track push with frontier AI
-
Industry
Robots can replace 74 percent of manual labor, but are not yet cheaper than humans
-
AI & Enterprise
Bill proposes youth \'shutdown\' feature for generative AI, letting parents set usage time
-
AI & Enterprise
Musk says retirement savings could become meaningless in 10 to 20 years
-
AI & Enterprise
OpenAI blocks attempts to extract protected reasoning, cites links to Moonshot AI
-
AI & Enterprise
Flood of legal AI tools leaves task of verifying fake case law
-
People
NHN Dooray wins personal information protection and utilisation technology award
-
AI & Enterprise
Adobe brings core creative tools to Gemini, expands features in Claude
-
AI & Enterprise
AWS expands availability of Anthropic\'s Claude in Seoul region
-
AI & Enterprise
Jensen Huang says using others\' AI models is competition, clashing with U.S. view of theft
-
AI & Enterprise
Liner launches \'Liner Actions MCP\' to connect 1,100 apps to AI agents
-
AI & Enterprise
Big Tech researchers urge policy oversight as AI automates AI research
-
AI & Enterprise
Weekley adds agents and personalisation to Owllo local AI, expands to hardware