Search results for Gemma 4
AI & Enterprise
Red Hat unveils Red Hat AI 3.5, strengthens security validation and GPU multitenancy
Red Hat introduced its new AI platform version, Red Hat AI 3.5, on Wednesday. The company said the release focuses on pre-deployment security validation, multitenancy for shared GPU infrastructure and improved observability, while also enhancing agent development. It launched EvalHub to assess models, RAG configurations and agents before deployment and generate compliance outputs. The model catalog adds more than 20 validated models and new scoring. Red Hat also highlighted new scheduling features and AutoRAG and Responses API updates.
AI & Enterprise
Nvidia begins mass production of Groq 3 LPX for AI agents
Nvidia has shifted its Groq 3 LPX inference accelerator for AI agents into mass production. The chip extends Nvidia\'s Vera Rubin data center platform and targets high-speed token generation to reduce decode latency in agent workflows. In an Artificial Analysis benchmark, it produced 3,400 tokens per second running the open-source Gemma 4 31B model with a 100,000-token context window. Nebius Group is among early adopters, and SpaceX was named a new flagship customer.
Industry
Nvidia unveils four new products ahead of earnings, focusing on token generation speed
Nvidia unveiled new AI inference accelerators and network infrastructure products at Hot Chips 2026 in the United States on Aug. 25, a day before its second-quarter earnings. The company said interactive inference accelerator Nvidia Grok 3 LPX has entered mass production, alongside Spectrum-X Multiplane Ethernet scaling and ScaleIn infrastructure acceleration technology. The announcements focused on inference, particularly token generation speed, and also highlighted inference speed, network scalability and agent-focused CPUs within the Vera Rubin platform.
-
AI & Enterprise
Google unveils Gemma 4 QAT for mobiles and laptops to cut AI memory use sharply
-
AI & Enterprise
Nvidia unveils 550 billion-parameter Nemotron 3 Ultra, starts mass production of Vera Rubin
-
AI & Enterprise
GitHub tool bypasses AI safeguards in some Meta, Google open-weight models, test finds
-
AI & Enterprise
Why Google Gemma 4 is in the spotlight as an open-source AI game changer
-
AI & Enterprise
LG unveils EXAONE 4.5 multimodal AI model focused on stronger reasoning
-
AI & Enterprise
Using Google\'s offline dictation app Eloquent, it polished speech in real time
-
AI & Enterprise
Evolving software pricing plans; Chinese AI firms revise strategy
-
AI & Enterprise
Google unveils open model Gemma 4, supports complex reasoning on low-power devices