Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Datajuicer Data juicer Custom Operator Development
- Workflow:Marker Inc Korea AutoRAG Data Creation Pipeline
- Workflow:Deepset ai Haystack RAG Pipeline
- Workflow:Interpretml Interpret EBM Model Export
- Workflow:Lucidrains X transformers Autoregressive Language Modeling
- Workflow:ContextualAI HALOs Model Evaluation
- Workflow:Recommenders team Recommenders Neural Collaborative Filtering
- Workflow:Tensorflow Tfjs Transfer Learning
- Workflow:MaterializeInc Materialize Integration Test Development
- Workflow:ClickHouse ClickHouse Server Deployment
Principles
- Principle:Axolotl ai cloud Axolotl Reference Model Setup
- Principle:Apache Airflow DAG Definition
- Principle:Ggml org Ggml BLAS Matrix Multiplication
- Principle:Ggml org Llama cpp Perplexity Computation
- Principle:Eventual Inc Daft Iceberg Catalog Creation
- Principle:DistrictDataLabs Yellowbrick Class Prediction Error Analysis
- Principle:Kubeflow Pipelines Iterative Training Termination
- Principle:CarperAI Trlx Distributed Logging
- Principle:Pytorch Serve Instance Segmentation
- Principle:Pyro ppl Pyro Online Statistics
Implementations
- Implementation:TA Lib Ta lib python Pip Install TA Lib
- Implementation:Openai Openai node Skills Resource
- Implementation:Open compass VLMEvalKit VisualGLM
- Implementation:Turboderp org Exllamav2 ExLlamaV2TokenizerHF
- Implementation:Apache Flink SplitsChange
- Implementation:Cypress io Cypress OTLPTraceExporter
- Implementation:Norrrrrrr lyn WAInjectBench LogisticRegression Fit
- Implementation:Wandb Weave Legacy Methods Lint Rules
- Implementation:Isaac sim IsaacGymEnvs AllegroKukaUtils
- Implementation:PeterL1n BackgroundMattingV2 OnnxRuntime InferenceSession
Heuristics
- Heuristic:Obss Sahi Class Agnostic vs Per Class NMS
- Heuristic:Evidentlyai Evidently Drift Detection Thresholds
- Heuristic:Intel Ipex llm Llama Padding Token Workaround
- Heuristic:ArroyoSystems Arroyo Stateful Operator TTL
- Heuristic:TobikoData Sqlmesh Forward Only Safety
- Heuristic:Spcl Graph of thoughts Scoring With Error Counting
- Heuristic:Fede1024 Rust rdkafka Transaction Error Recovery
- Heuristic:Cohere ai Cohere python Warning Deprecated Legacy Generate API
- Heuristic:Google research Deduplicate text datasets Ulimit File Descriptors For Merge
- Heuristic:Speechbrain Speechbrain Score Normalization Tips
Environments
- Environment:OpenRLHF OpenRLHF Flash Attention Environment
- Environment:FMInference FlexLLMGen HuggingFace Access
- Environment:Mlc ai Mlc llm Metal macOS iOS Environment
- Environment:NVIDIA DALI CUDA GPU Environment
- Environment:Alibaba ROLL Ascend NPU Environment
- Environment:Pola rs Polars GPU Execution Environment
- Environment:Vllm project Vllm GitHub
- Environment:OpenGVLab InternVL PyTorch CUDA
- Environment:Datahub project Datahub Docker Quickstart Environment
- Environment:Run llama Llama index OpenAI API Configuration