AI Launch Radar

Cloud AI infra & ops automation race — edge/inference optimizations

Cloud AI infra & ops automation race — edge/inference optimizations

Key Questions

What optimizations are occurring in cloud AI infrastructure?

The highlight covers edge and inference optimizations such as HF browser 8B at 60t/s, Gemma 4 on iPhone/Jetson, and local tools like OpenMed. Companies including Nutanix, Cloudflare, and Mistral Forge are advancing these capabilities.

What is the significance of Etched's valuation?

Etched, an AI chip startup founded by Harvard dropouts in 2022, reached a $10.3B valuation after closing a $300M round. This reflects growing investor interest in specialized hardware for AI inference and edge computing.

How are local and edge compute trends evolving?

Local/edge compute backing is growing with tools like Megaphone for on-device Mac dictation and sllm priced at $5-10/mo. This supports broader adoption of efficient AI models on consumer and edge devices.

HF browser 8B 60t/s; VueBuds wearable vision; Gemma 4 iPhone/Jetson; OpenMed local; GitHub strains; Qwen Cloud; Nutanix K8s; sllm $5-10/mo; Aiden; Cloudflare Think; Cognichip/Rebellions/ScaleOps/Mistral Forge/Together/PrismML/Nscale/Gimlet/MatX/Tinybox. Local/edge compute backing grows.

Sources (2)
Updated Jul 23, 2026
What optimizations are occurring in cloud AI infrastructure? - AI Launch Radar | NBot | nbot.ai