Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Vllm project Vllm Speculative Decoding
- Workflow:PrefectHQ Prefect Per Worker Task Concurrency
- Workflow:Predibase Lorax Structured JSON Output
- Workflow:LLMBook zh LLMBook zh github io LoRA Finetuning
- Workflow:OWASP Www project top 10 for large language model applications GenAI Red Team Testing
- Workflow:Microsoft Agent framework Multi Agent Sequential Orchestration
- Workflow:Apache Druid SQL Based Data Ingestion
- Workflow:Junyanz Pytorch CycleGAN and pix2pix Pretrained Inference
- Workflow:Huggingface Transformers Model Quantization
- Workflow:Mlc ai Web llm Basic Chat Completion
Principles
- Principle:Openai Whisper Cross Attention Alignment
- Principle:OWASP Www project top 10 for large language model applications Risk Scoring and Prioritization
- Principle:Microsoft Autogen Swarm Orchestration
- Principle:BerriAI Litellm Cache Initialization
- Principle:Rapidsai Cuml Classification Evaluation
- Principle:Apache Airflow DAG Deployment
- Principle:Facebookresearch Audiocraft Compression Training Execution
- Principle:Haotian liu LLaVA Checkpoint Extraction
- Principle:Interpretml Interpret Interactive Visualization Rendering
- Principle:Promptfoo Promptfoo Test Suite Construction
Implementations
- Implementation:NVIDIA NeMo Curator ClientPartitioning
- Implementation:Diagram of thought Diagram of thought Node Edge Status Protocol Config
- Implementation:TobikoData Sqlmesh Context Render Evaluate Fetchdf
- Implementation:Microsoft Playwright BrowserDispatcher
- Implementation:CARLA simulator Carla Command DestroyActor
- Implementation:Openai Openai agents python Run Examples
- Implementation:Google deepmind Mujoco Engine Print
- Implementation:Neuml Txtai RAG Prompt Template
- Implementation:Deepseek ai Janus ODE Denoising Loop
- Implementation:Run llama Llama index MockLLM
Heuristics
- Heuristic:Astronomer Astronomer cosmos Deprecation Migration Paths
- Heuristic:Scikit learn contrib Imbalanced learn Sampling Strategy Selection
- Heuristic:Helicone Helicone ClickHouse ReplacingMergeTree FINAL
- Heuristic:Unstructured IO Unstructured Strategy Fallback Chain
- Heuristic:Shiyu coder Kronos Two Stage Finetuning Strategy
- Heuristic:MaterializeInc Materialize Docker Image Cache Lookup
- Heuristic:Microsoft Onnxruntime Flash Attention Optimization
- Heuristic:Apache Airflow Task Idempotency Pattern
- Heuristic:Datajuicer Data juicer Checkpoint Resumption Strategy
- Heuristic:Explodinggradients Ragas Embedding Batch Size Tuning
Environments
- Environment:Facebookresearch Habitat lab SLURM Distributed Environment
- Environment:Lm sys FastChat API Keys And Credentials
- Environment:Huggingface Trl vLLM Generation Environment
- Environment:Tensorflow Serving Docker Runtime Environment
- Environment:Testtimescaling Testtimescaling github io GitHub Actions Runner
- Environment:Allenai Open instruct CUDA GPU Training
- Environment:Huggingface Alignment handbook DeepSpeed Multi Node
- Environment:ArroyoSystems Arroyo Webui Runtime
- Environment:Google research Deduplicate text datasets Python HuggingFace Environment
- Environment:Google research Deduplicate text datasets Python TFDS Environment