Search the public knowledge base.
22 results

Imagine an LLM solving a math problem. It samples a full reasoning trace, you run a verifier, and the verifier returns a number: per...

A diagram-rich generated explanation from the public library.

A diagram-rich generated explanation from the public library.

A diagram-rich generated explanation from the public library.

A diagram-rich generated explanation from the public library.

A diagram-rich generated explanation from the public library.

A diagram-rich generated explanation from the public library.

A diagram-rich generated explanation from the public library.

KV Caching

A diagram-rich generated explanation from the public library.

A diagram-rich generated explanation from the public library.

VLA (Vision Language Action Models)

MoE architecture for efficient LLM scaling via specialized experts

RL framework that approximates maximum likelihood for binary-outcome tasks.

GRPO and RL for LLM's

Stable JEPA-based world model that learns and plans from raw pixels.

Language models that recursively refine or compose intermediate reasoning/representations.

Mamba-3: Improved Sequence Modeling using State Space Principles

Transformers

Diffusion and flow-matching

LORA Fine-tuning (Low rank adaption)

policy gradient methods