Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Ray project Ray Build and Release Pipeline
- Workflow:Ollama Ollama Custom Model Creation
- Workflow:Huggingface Peft Seq2Seq AdaLoRA Finetuning
- Workflow:DistrictDataLabs Yellowbrick Regression Model Evaluation
- Workflow:Langchain ai Langgraph Functional API Workflow
- Workflow:OpenRLHF OpenRLHF Iterative DPO
- Workflow:Farama Foundation Gymnasium Vectorized Environment Training
- Workflow:Astronomer Astronomer cosmos Kubernetes dbt execution
- Workflow:ChenghaoMou Text dedup Suffix Array Deduplication
- Workflow:Microsoft Semantic kernel Plugin Integration And Function Calling
Principles
- Principle:Cohere ai Cohere python Embed Response Processing
- Principle:Microsoft Playwright Run Tests with Tracing
- Principle:Langgenius Dify Retrieval Configuration
- Principle:Online ml River Online Linear Regression
- Principle:Scikit learn contrib Imbalanced learn Sampler Compatibility Checking
- Principle:CrewAIInc CrewAI Output Processing
- Principle:Huggingface Datatrove Disk Writing Framework
- Principle:SeldonIO Seldon core Pipeline Readiness Verification
- Principle:BerriAI Litellm Database Setup
- Principle:Mbzuai oryx Awesome LLM Post training Multi Sheet Excel Export
Implementations
- Implementation:SeleniumHQ Selenium DevTools AddListener
- Implementation:Hiyouga LLaMA Factory Tuner
- Implementation:Mlc ai Web llm Create Service Worker MLC Engine
- Implementation:Nautechsystems Nautilus trader OrderFactory Market Limit
- Implementation:Scikit learn Scikit learn FetchLfw
- Implementation:Evidentlyai Evidently Guardrails Trace
- Implementation:Scikit learn Scikit learn OPTICS
- Implementation:Open compass VLMEvalKit MiniCPM V
- Implementation:Infiniflow Ragflow Admin Whitelist Page
- Implementation:Speechbrain Speechbrain Prepare Voicebank MTL
Heuristics
- Heuristic:NVIDIA NeMo Aligner Higher Stability Log Probs
- Heuristic:Arize ai Phoenix Warning Deprecated VertexAIModel
- Heuristic:Romsto Speculative Decoding Seed Fixing For Reproducibility
- Heuristic:Alibaba MNN NC4HW4 Data Layout
- Heuristic:Astronomer Astronomer cosmos Cache Strategy Optimization
- Heuristic:VainF Torch Pruning Pruning Ratio vs Parameter Ratio
- Heuristic:Kubeflow Pipelines Cache Staleness In Recursive Pipelines
- Heuristic:SqueezeAILab ETS Thread Parallelism Suppression
- Heuristic:Microsoft Onnxruntime Threading Configuration Tips
- Heuristic:ContextualAI HALOs LoRA Merge At Save
Environments
- Environment:Marker Inc Korea AutoRAG API Keys And Credentials
- Environment:Apache Shardingsphere Java Runtime Environment
- Environment:DataExpert io Data engineer handbook Statsig API Environment
- Environment:PrefectHQ Prefect AI Integration Credentials
- Environment:Sgl project Sglang CUDA Runtime
- Environment:Huggingface Datasets Lance Dependencies
- Environment:Haosulab ManiSkill Motion Planning Deps
- Environment:Allenai Open instruct vLLM Inference
- Environment:LLMBook zh LLMBook zh github io Data Processing Environment
- Environment:Allenai Open instruct CUDA GPU Training