Weekly AI Research Digest

The Field, Week of July 13, 2026

Compiled for History's Future: The Singularity Is Here — curated from Hugging Face trending models, datasets, and research papers.

Field Pulse

The week of July 13, 2026 is defined by a single, striking development: extreme model compression has gone from experiment to infrastructure. Prism-ML's Bonsai-27B family — one variant at 1-bit and another at a novel 2-bit ternary scheme — are trending with scores above 800 and 500 respectively, dwarfing most other releases. These aren't fringe research artifacts. They run Qwen3.6-27B-class reasoning on hardware that would previously have struggled with a 7B model. The singularity thesis has always predicted that intelligence would become cheap before it became visible; the Bonsai family is that prediction arriving early.

Document intelligence is the field's other pressure point this week. Baidu's Unlimited-OCR continues to trend at massive scale (Score 441, 1.1M downloads), while ATH-MaaS's OvisOCR2 arrived fresh on July 13 with multimodal OCR improvements. Meanwhile, Wan-AI's Wan-Dancer-14B lands as the week's surprise: a 14B music-to-dance and image-to-video generation model, signaling that video synthesis — long the province of massive closed systems — is now tractable in open weights. Taken together, the week's outputs compress two years of expected progress into a single fortnight.

Thematic Overview

On-Device Compression Document & Vision AI Multimodal Generation Bonsai-27B (1-bit GGUF) Ternary-Bonsai-27B (2-bit) MiniCPM5-1B Thinking GGUF ThinkingCap Qwen3.6-27B K-Quantization study (paper) Baidu Unlimited-OCR ATH-MaaS OvisOCR2 Hy3-GGUF (Tencent Hy3) LittleBit sub-1-bit (paper) SenseNova Vision Corpus 50M Wan-Dancer-14B (video) MOSS-Transcribe-Diarize SupraLabs Reasoning Corpus Antidoom-Mix (LiquidAI) NVIDIA Open-SWE-Traces SINGULARITY THESIS History's Future · ashokmehan.com
On-Device Compression Document & Vision AI Multimodal Generation

Top Trending Models

Ternary-Bonsai-27B-GGUF
Text Generation · 2-bit Ternary GGUF
On-Device Compression
Score: 812
Prism-ML's 2-bit ternary quantization of Qwen3.6-27B — the highest-trending model this week by a wide margin. Runs a 27B reasoning model in ternary weight space, dramatically reducing memory while preserving benchmark performance.
View on HF →
Bonsai-27B-GGUF
Text Generation · 1-bit GGUF
On-Device Compression
Score: 521
Sibling to Ternary-Bonsai — a full 1-bit quantization of Qwen3.6-27B. Together the Bonsai family makes 27B-class reasoning accessible on consumer-grade hardware without GPU servers.
View on HF →
Baidu Unlimited-OCR
Image-Text-to-Text · VLM
Document & Vision AI
⬇ 1.1M♥ 1,773Score: 441
Baidu's multilingual VLM OCR continues its dominance — now among the most-downloaded vision models on HF. Universal document understanding across scripts and languages at production scale.
View on HF →
ThinkingCap Qwen3.6-27B
Text Generation · Token-Efficient Reasoning
On-Device Compression
Score: 162
BottlecapAI's approach to making Qwen3.6-27B thinking more token-efficient — reduces chain-of-thought verbosity while maintaining reasoning accuracy, cutting inference cost for reasoning tasks.
View on HF →
OvisOCR2
Image-Text-to-Text · OCR VLM
Document & Vision AI
Score: 161
ATH-MaaS's second-generation multimodal OCR model, released July 13. Improves on structured document parsing and mixed text-visual layouts compared to the original OvisOCR.
View on HF →
Hy3-GGUF
Text Generation · MoE GGUF
Document & Vision AI
Score: 148
AngelSlim's GGUF quantization pack of Tencent's Hunyuan-3 MoE. The full Hy3 model launched last week; quantized variants arrived July 13, making the architecture accessible on consumer hardware.
View on HF →
MOSS-Transcribe-Diarize
Audio · ASR with Speaker Diarization
Multimodal Generation
Score: 128
OpenMOSS Team's automatic speech recognition model with native speaker diarization — identifies both what was said and who said it, enabling meeting transcription and conversation analytics at no cost.
View on HF →
MiniCPM5-1B Thinking GGUF
Text Generation · Ultra-Compact Reasoning
On-Device Compression
Score: 129
A 1-billion-parameter model with chain-of-thought reasoning, fine-tuned on Claude Opus and Fable 5, released July 13. Represents the extreme end of the compression curve — reasoning at 1B parameters.
View on HF →
Wan-Dancer-14B
Video Generation · Image/Music to Dance
Multimodal Generation
Score: 108
Wan-AI's 14B video generation model (released July 10) converts reference images and music inputs into animated dance videos. Open-weight video synthesis is arriving faster than anticipated.
View on HF →

Notable Datasets

This Week's Papers