Tag: inference
Chinese Startup DeepSeek Develops AI Chip for Inference Acceleration
Chinese AI startup DeepSeek is quietly developing a dedicated chip to enhance inference performance, consulting partners and recruiting specialists.
Read MoreOpenAI Unveils Jalapeno Chip to Lower ChatGPT Operating Costs
OpenAI introduces Jalapeno, its first specialized chip designed to reduce resource use for running AI models like ChatGPT by up to 50 percent.
Read MoreAsus Unveils ExpertCenter Pro ET900N G3 Desktop with Nvidia GB300 for AI Workloads
Asus launches ExpertCenter Pro ET900N G3, a desktop powered by Nvidia GB300 designed for AI training and inference with 748GB memory capacity.
Read MoreIntel Unveils Crescent Island AI Accelerator With Up to 480 GB LPDDR5X Memory
Intel introduces Crescent Island, a high-performance AI accelerator designed for inference tasks, featuring up to 480 GB LPDDR5X memory and 350 W power draw.
Read MoreAdvances in AI Hardware Could Lower Inference Costs, But Consumer Prices May Remain Stable
New AI processors promise cheaper inference, yet rising infrastructure expenses keep consumer costs largely unchanged for now.
Read MoreHarnessing Ambient Noise: Thermodynamic Computing Offers Energy-Efficient AI Processing
Thermodynamic computing proposes using environmental noise as a resource to reduce energy spent on AI model training and inference.
Read MoreAlphabet Collaborates with Marvell on New AI Inference Chips
Alphabet and Marvell are negotiating to develop two specialized AI chips aimed at enhancing inference performance and data transfer speeds.
Read MoreNVIDIA Introduces Groq 3 LPU: A Shift Toward Deterministic AI Inference
NVIDIA’s Groq 3 LPUs mark a departure from traditional AI accelerators, enabling deterministic inference for next-gen heterogeneous AI platforms.
Read MoreNvidia Develops Tailored Groq AI Chips for Chinese Market Amid Export Controls
Nvidia is preparing a customized version of Groq AI chips for China, addressing U.S. export restrictions while enhancing inference performance.
Read MoreNvidia Unveils Groq 3 LPU Chip to Boost AI Inference Performance
Nvidia introduces the Groq 3 LPU, a new inference accelerator chip designed to enhance token-level processing with high throughput and low latency.
Read MoreNvidia Highlights Significant AI Inference Cost Reduction with Blackwell Architecture
Nvidia reports that the Blackwell AI architecture has cut neural network inference costs by up to 10 times, leveraging both hardware and software advancements.
Read MoreNvidia CEO Clarifies Commitment to OpenAI Investment Amid Rumors
Nvidia’s Jensen Huang affirms ongoing plans to invest in OpenAI despite rumors of lost interest following hardware alternatives.
Read MoreRecent Posts
- Memory Chip Makers Face Potential DRAM Oversupply by 2028 Amid Massive Capacity Expansions
- Samsung’s Galaxy Z Fold8 Expected to Lead New Foldable Smartphone Market
- Apple Explored Intel Chips as Backup for Mac Pro Before Discontinuation
- SpaceX Reschedules Starship V3 Test Launch for Thursday
- SK hynix Considers Building Memory Chip Factory in the U.S. to Address High Prices