Install
Stay on top of the latest research, models, and breakthroughs in AI. Read by 150,000+ engineers, data scientist, and researchers.
- 18articles · 30d
- 17+ hour agolatest article
- Aug 15, 2026earliest in window
- 78%with images
- 290avg words
- Science & Technology 17
- Computers & Electronics 13
- Software Dev. 13
- Science & Nature 3
- News 2
- Business & Industrial 1
- Economy, Business & Finance 1
- Education & Jobs 1
Please confirm you are human
This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.
A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.
News
Edge0 Streams an 8B AI Model From SSD Using Only 1 GiB
22+ hour, 23+ min ago (262+ words) Edge0 releases an 8B sparse MoE that runs in under 1 GiB of active memory on Apple Silicon, streaming experts from SSD on demand. Edge0 has released the 8B checkpoint and an accompanying streaming runtime that load mixture-of-experts weights from SSD as each token needs…...
Zed 1.19 Ships Call Hierarchy and Live Search to Rival Top Editors
4+ day, 14+ hour ago (313+ words) Zed's weekly release adds call hierarchy navigation, multi-select Git staging, auto language detection for scratch buffers, and live search-as-you-type by default. Zed 1.19 is out, and it's a polish release rather than a headline-feature swing. The update adds call hierarchy navigation,…...
Liquid AI's Pipette Exposes What Server Benchmarks Hide About Phone AI
2+ week, 6+ day ago (290+ words) Liquid AI and Artificial Analysis release an open-source suite that measures model quality, speed, latency, and memory across real phones, laptops, and embedded hardware. Model cards keep telling you what a language model can do on an H100. They rarely tell…...
vLLM | AI Companies
4+ week, 1+ day ago (139+ words) Open-source LLM inference and serving engine, originated at UC Berkeley's Sky Computing Lab. Built around PagedAttention for efficient KV cache memory management, with continuous batching, tensor and pipeline parallelism, and quantization support (FP8, GPTQ, AWQ). Supports 200+ Hugging Face model architectures with…...