Reading up on Qwen3.5
2 deep · digging since mar 03
- Running LLMs in the Browser with Three.js
Three-LLM enables running GPT-2, SmolLM2, Qwen, and Phi language models in the browser by converting their inference graphs into Three.js TSL compute shaders on WebGPU.
- Alibaba's small, open source Qwen3.5-9B beats OpenAI's gpt-oss-120B and can run on standard laptops
Alibaba's Qwen3.5-9B outperforms OpenAI's 120B-parameter gpt-oss on benchmarks while being 13 times smaller and capable of running on standard laptops.