Search results for AI inference
AI & Enterprise
Baseten joins OpenAI marketplace, open models available directly in Codex
AI inference infrastructure firm Baseten said on Sept. 29 it is working with OpenAI to provide open models through OpenAI\'s enterprise marketplace. Developers can select and use Baseten open models directly within OpenAI\'s Codex coding tool. OpenAI enterprise customers can use Baseten models under their existing OpenAI spending commitments, either in Codex or via the Responses API. Baseten said it runs inference on U.S. infrastructure and does not store prompts.
Industry
Distributed AI gains attention as response tool for workforce shortages in industry
Distributed AI is emerging as a key option for industrial sites facing growing labour shortages, as robots and equipment increasingly need to make on-the-spot decisions. Access Partnership said autonomous industrial operations could generate more than $23 billion a year in economic value by 2035, and companion and assistive robots more than $5.9 billion. The report also put total annual value across industry, care and export devices at over $50 billion by 2035.
Industry
Semifive signs 70.3 billion won contract to develop AI inference accelerator with U.S. fabless firm
Semifive said on Monday it signed a contract with a U.S. AI fabless company to develop a next-generation AI inference accelerator. The deal is worth about 70.3 billion won ($52 million) and marks its first spec handoff project secured in the North American market. The company said it was its largest single contract and outlined plans for tape-out in the first half of 2027 and mass production from 2028.
-
Industry
Qualcomm chip-based Microsoft Surface PCs show 80 percent faster local AI inference
-
Industry
Snapdragon Summit 2026: Qualcomm expands Snapdragon X2 Linux support to broaden agentic AI development base
-
Industry
Quality and context are hurdles to spread of on-device agents
-
Industry
Qualcomm fleshes out \'AI smartphone\' vision, up to 1 million tokens a day
-
AI & Enterprise
Delos Data raises over $100 million, unveils network chip for AI inference
-
AI & Enterprise
South Korea to raise ICT R&D budget to 2 trillion won, focus on three mega projects
-
Industry
Qualcomm boosts Hexagon NPU memory 50 percent, targets AI agents
-
Industry
FuriosaAI sets up Singapore unit, steps up push into Asia-Pacific market
-
AI & Enterprise
Lightbits launches KV cache engine to boost GPU performance
-
Industry
WD: 95 percent of companies say data value rises after adopting AI
-
AI & Enterprise
OpenAI aims to expand ChatGPT ads, targeting $1 billion a year
-
AI & Enterprise
OpenAI prepares inference residency in South Korea, in talks with local partners to secure GPUs
-
AI & Enterprise
HS Hyosung Information System targets AI inference infrastructure market with Arm server GreenCore and NPU
-
AI & Enterprise
Qualcomm teams up with AWS to build AI infrastructure for inference, shares surge
-
Industry
Semifive begins mass production of Samsung 4-nanometer AI inference accelerator