Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Ray project Ray Build and Release Pipeline
- Workflow:SqueezeAILab ETS Answer Evaluation Pipeline
- Workflow:Scikit learn Scikit learn Ensemble Model Building
- Workflow:Predibase Lorax Multi Adapter Merging
- Workflow:Lucidrains X transformers Non Autoregressive Masked Generation
- Workflow:TA Lib Ta lib python Abstract API Usage
- Workflow:Apache Spark Release Process
- Workflow:Openai Openai node Audio Processing
- Workflow:ARISE Initiative Robomimic Trained Policy Evaluation
- Workflow:ARISE Initiative Robosuite Environment Setup And Simulation
Principles
- Principle:Vespa engine Vespa CPP Compilation
- Principle:Alibaba MNN PyMNN Installation
- Principle:Apache Hudi Table Schema Definition
- Principle:Speechbrain Speechbrain TTS Inference Pipeline
- Principle:Danijar Dreamerv3 Agent Initialization
- Principle:ArroyoSystems Arroyo Barrier Propagation
- Principle:Alibaba ROLL RLVR Configuration
- Principle:Wandb Weave Version Bump Release
- Principle:FlowiseAI Flowise Lead Capture
- Principle:Datahub project Datahub Emitter Instantiation
Implementations
- Implementation:Openai Whisper BasicTextNormalizer
- Implementation:Huggingface Datasets Dataset Num Rows
- Implementation:Webdriverio Webdriverio Error Handling Pattern
- Implementation:Huggingface Datasets Dataset Shuffle
- Implementation:Mlflow Mlflow Log Artifact
- Implementation:Open compass VLMEvalKit DREAM
- Implementation:Lucidrains X transformers DPO Forward
- Implementation:Datajuicer Data juicer TextFormatter
- Implementation:Apache Druid Hjson Context
- Implementation:Apache Paimon FunctionDefinition
Heuristics
- Heuristic:Apache Airflow Variable Access Pattern
- Heuristic:Apache Airflow Task Dependency Isolation
- Heuristic:Vespa engine Vespa Log Level Inheritance Polling
- Heuristic:Bentoml BentoML Thread Env Vars Setting
- Heuristic:Ggml org Llama cpp Thread Count Tuning
- Heuristic:Mit han lab Llm awq Kernel Selection Thresholds
- Heuristic:Huggingface Datasets Warning Deprecated Pandas Builder
- Heuristic:Tencent Ncnn Lightmode Memory Optimization
- Heuristic:Recommenders team Recommenders SAR Cold Start Items
- Heuristic:Pyro ppl Pyro MCMC Warmup Adaptation
Environments
- Environment:BerriAI Litellm Provider API Credentials
- Environment:Zai org CogVideo SAT Framework Environment
- Environment:Volcengine Verl CUDA GPU Environment
- Environment:Openai Whisper Triton
- Environment:Apache Airflow Database Backend Environment
- Environment:Unslothai Unsloth CUDA VLLM
- Environment:Ggml org Llama cpp Vulkan GPU Environment
- Environment:Run llama Llama index Fsspec Remote Storage
- Environment:Ggml org Llama cpp Metal GPU Environment
- Environment:Microsoft Onnxruntime Nodejs Runtime Environment