Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:NVIDIA TransformerEngine Accelerate HF Llama With TE
- Workflow:Kornia Kornia Edge Detection Pipeline
- Workflow:Intel Ipex llm RAG With LangChain
- Workflow:Interpretml Interpret EBM Model Export
- Workflow:Puppeteer Puppeteer Browser Installation And Management
- Workflow:Infiniflow Ragflow Knowledge Base Document Ingestion
- Workflow:Huggingface Alignment handbook QLoRA Single GPU Finetuning
- Workflow:Microsoft Onnxruntime ORTModule Training
- Workflow:Speechbrain Speechbrain Whisper ASR Finetuning
- Workflow:Predibase Lorax Single LoRA Inference
Principles
- Principle:Mit han lab Llm awq W8A8 Vision Encoder Quantization
- Principle:Triton inference server Server Classification Postprocessing
- Principle:Predibase Lorax Inference Request Construction
- Principle:InternLM Lmdeploy SmoothQuant Quantization
- Principle:Ggml org Llama cpp HF to GGUF Conversion
- Principle:Isaac sim IsaacGymEnvs Task Registration
- Principle:FlowiseAI Flowise Dynamic Form Input
- Principle:Pyro ppl Pyro Causal Effect Estimation
- Principle:Pyro ppl Pyro Custom Distribution Framework
- Principle:Microsoft Onnxruntime TensorBoard Monitoring
Implementations
- Implementation:Sgl project Sglang Marlin MoE WNA16 Template
- Implementation:Predibase Lorax Prepare Chat Input
- Implementation:NVIDIA TransformerEngine TELlamaDecoderLayer
- Implementation:LMCache LMCache LMCacheControllerManager Init
- Implementation:Mlc ai Mlc llm Config Base
- Implementation:Microsoft DeepSpeedExamples Add Argument CIFAR
- Implementation:Onnx Onnx Model Subgraph Extractor
- Implementation:Alibaba MNN FlatBuffers IDL Gen Text
- Implementation:Openai Openai python Fine Tuning Job Cancelled Webhook
- Implementation:Predibase Lorax GPTQ Quant Linear
Heuristics
- Heuristic:ThreeSR Awesome Inference Time Scaling API Rate Limiting Tip
- Heuristic:Alibaba MNN Memory Mode Selection
- Heuristic:Hpcaitech ColossalAI Flash Attention Dtype Restriction
- Heuristic:Junyanz Pytorch CycleGAN and pix2pix Batch Size One Default
- Heuristic:Heibaiying BigData Notes Spark Streaming Local Threads Tip
- Heuristic:Triton inference server Server Documentation Standards
- Heuristic:Haosulab ManiSkill Physics Solver Tuning
- Heuristic:VainF Torch Pruning Over Pruning Prevention
- Heuristic:Vllm project Vllm GPU Memory Utilization Tuning
- Heuristic:AUTOMATIC1111 Stable diffusion webui NaN Detection And Precision Fixes
Environments
- Environment:Mistralai Client python Realtime Transcription Environment
- Environment:Infiniflow Ragflow GPU CUDA Environment
- Environment:Datahub project Datahub Frontend Build
- Environment:Confident ai Deepeval LLM Provider Credentials
- Environment:Ggml org Llama cpp CUDA GPU Environment
- Environment:Langchain ai Langchain Unit Test Network Isolation
- Environment:Google deepmind Dm control EGL Headless Rendering
- Environment:Turboderp org Exllamav2 Flash Attention Backend
- Environment:Infiniflow Ragflow Docker Infrastructure
- Environment:BerriAI Litellm Python Runtime