Search results for KV cache
AI & Enterprise
Lightbits launches KV cache engine to boost GPU performance
Lightbits Labs has launched its AI inference software engine, Infera, aimed at improving inference performance and cost efficiency. The company says Infera expands KV cache data beyond limited high-bandwidth memory attached to GPUs and targets large language model operations with long context windows or many concurrent sessions. It manages KV cache across GPU memory, DRAM and NVMe storage using predictive prefetching and supplies context in small blocks.
AI & Enterprise
Agent era accelerates multi-model use; small models set to take larger share
Multi-model use inside generative AI services is already common, with different-sized models dividing work, a Ravlub researcher said. Small models handle routine tasks such as search, information gathering and summarisation, while frontier models take key decisions and generate outputs. As agents spread, repeatedly calling top-performance models can push inference costs higher. He said small models will not replace frontier models, citing scaling, but role-based use of multiple model sizes will expand.
Industry
SK hynix to invest 54 trillion won in Yongin, Cheongju to expand DRAM and NAND fabs
SK hynix will build two new chip fabrication plants in Yongin and Cheongju and invest about 54 trillion won, after a board resolution on Aug. 7. It plans 35.2 trillion won for Yongin Y2 and 19.1 trillion won for Cheongju M17. Yongin will be a DRAM base and Cheongju a NAND base. The company said it decided after closely reviewing demand and will phase spending through 2031.
-
AI & Enterprise
OpenAI says GPT-5.6 Sol optimises GPU efficiency itself, cuts inference costs 20 percent
-
Industry
SK Hynix wraps up LTA talks with about 10 customers including core clients, negotiating 2027 volumes and prices
-
Industry
Samsung Electronics, SK Hynix DRAM operating margin nears 80 percent as memory boom holds firm
-
AI & Enterprise
Neocloud market shifts as SoftBank and Meta enter
-
AI & Enterprise
Dinotisia unveils KV cache compression technology to ease AI computing bottlenecks
-
Industry
AI demand debate clouds memory order visibility
-
AI & Enterprise
DeepSeek keeps prices 75% lower with GPT-level performance; memory efficiency seen as key
-
Industry
Memory shortage shadow looms over Computex 2026 as HBM4 supply race intensifies
-
Industry
QLC SSD demand rises as HDD supply tightens on cold data surge
-
Industry
SK Hynix earnings surprise likely as NAND exports surge
-
Industry
Market uneasy despite outlook for record results at Samsung Electronics, SK Hynix
-
Industry
Dnotitia unveils AI storage strategy, aims to move beyond simple storage
-
AI & Enterprise
NC AI unveils industry-focused AI model VAETKI, cuts memory use 83 percent