🚨 Meta's 4-bit 30B in 24 GB of VRAM, NVIDIA's 448 ms full-duplex turns, and an agent runtime that forks 5× faster than Docker
Six releases, and not one of them was decided by a benchmark. VRAM ceilings, license text, and deployment targets did the deciding. Meta got a 30B under 20 GB. webAI shipped 1.78 GiB you may not sell. NVIDIA published permissive weights and tagged them research-only.