#AI inference
7 stories taggedAI inference.

d-Matrix Plugs Its AI Chips Into NVIDIA's Factory Platform
A chipmaker building specialist inference processors just found a shortcut to deploying them at scale, and the technical gap it closes is bigger than the announcement looks.

The Real AI Bottleneck Is Not the Chip. It Is Where the Data Lives.
AI systems are straining under real-time demands, and the weak point is not processing power. It is memory, storage, and the pipes connecting them.

Your Gaming PC Could Earn Money Running AI Jobs While You Sleep
A handful of start-ups now pay ordinary people to share spare computing power from home computers, gaming rigs, and basement servers. Here is how it works and what to watch for.

OpenAI construyó su propio chip de IA. Esto es lo que significa para tu cartera y tus tiempos de espera.
El silicio personalizado de OpenAI, llamado Jalapeño, promete respuestas más rápidas y facturas de energía más bajas. Los volúmenes pequeños se enviarán antes de finales de 2025, con un impulso mayor en 2027.

Nvidia's simple algebra trick makes AI model handoffs up to 25 times faster
Swapping between AI models mid-task is brutally expensive. Nvidia researchers found that a basic linear equation can do what previously needed a full restart, cutting costs and wait times sharply.

Starcloud raises $250 million to build AI data centres in space
The satellite startup is now valued at $2.3 billion and is betting that SpaceX's next-generation rocket will make processing data in orbit genuinely competitive with servers on the ground.

La startup francesa Kog afirma que sus GPUs existentes pueden ejecutar IA diez veces más rápido. Así es cómo.
Kog sostiene que solo el software puede acelerar dramáticamente la inferencia de IA en los chips que ya poseen las empresas. La demostración de septiembre será la prueba definitiva.