Billion-parameter LLMs now run in real time on phones due to advances in model compression, quantization, and efficient architectures, not just faster chips.
About HuggingFace · Meta · DeepSeek-R1 · AWQ · ExecuTorch · Llama.cpp Filed #ai #ai-research #developer-tools #machine-learning #mobile Related As Rocks May Think Coding agents combined with reasoning LLMs have become automated scientists, enabling a golden age where all computer science problems appear tractable through vast inference compute.
also on DeepSeek-R1 , #ai-research The Future is for Everyone Mark Zuckerberg argues that distributing superintelligence broadly to individuals, rather than centralizing it, will empower people, drive invention, and ensure a balanced, prosperous AI future through personal agency and widespread access.
also on HuggingFace , Meta , #ai , #developer-tools