Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:DataExpert io Data engineer handbook Dimensional Data Modeling Environment Setup
- Workflow:Google deepmind Dm control MJCF Model Composition
- Workflow:Spotify Luigi Database Ingestion Pipeline
- Workflow:Dagster io Dagster ML Pipeline
- Workflow:Confident ai Deepeval LLM Tracing and Observability
- Workflow:TobikoData Sqlmesh Plan and apply deployment
- Workflow:Kubeflow Pipelines Iterative Model Training
- Workflow:FlagOpen FlagEmbedding Benchmark Evaluation
- Workflow:Huggingface Transformers PEFT Adapter Integration
- Workflow:Apache Paimon Data Ingestion With Ray Sink
Principles
- Principle:Hiyouga LLaMA Factory Rotary Position Embedding
- Principle:Heibaiying BigData Notes Spark Data Writing
- Principle:Predibase Lorax Chat Template Rendering
- Principle:Ollama Ollama GGUF Model Conversion GptOss
- Principle:Googleapis Python genai Image Result Processing
- Principle:DistrictDataLabs Yellowbrick Discrimination Threshold Analysis
- Principle:Huggingface Datatrove Duplicate Clustering
- Principle:Openclaw Openclaw Access Policy Configuration
- Principle:Dagster io Dagster Materialization Metadata
- Principle:Zai org CogVideo Activation Normalization
Implementations
- Implementation:Haosulab ManiSkill PickSingleYCB
- Implementation:Tencent Ncnn ParamDict
- Implementation:Online ml River Imblearn HardSampling
- Implementation:Duckdb Duckdb HyperLogLog
- Implementation:Elevenlabs Elevenlabs python ProjectExternalAudioResponseModel
- Implementation:CARLA simulator Carla Python Commands Bindings
- Implementation:Tensorflow Serving Tfrt Saved Model Factory
- Implementation:Mit han lab Llm awq CLIPVisionTower
- Implementation:SeleniumHQ Selenium DevTools Event
- Implementation:Run llama Llama index GraphStore Types
Heuristics
- Heuristic:Scikit learn Scikit learn Random State Management
- Heuristic:Haotian liu LLaVA Quantization MM Projector Exclusion
- Heuristic:Langgenius Dify SQL Escape Backslash First
- Heuristic:Bigscience workshop Petals NF4 Quantization Default On CUDA
- Heuristic:Deepspeedai DeepSpeed FP16 Convergence Tips
- Heuristic:InternLM Lmdeploy KV Cache Memory Tuning
- Heuristic:Huggingface Trl Disable Dropout For RL Training
- Heuristic:Testtimescaling Testtimescaling github io Hardcoded IDs vs Registry
- Heuristic:PrefectHQ Prefect SQLite Performance Tuning
- Heuristic:Iamhankai Forest of Thought Tree Iteration Scaling
Environments
- Environment:AUTOMATIC1111 Stable diffusion webui GPU Compute Backend
- Environment:FlagOpen FlagEmbedding Finetuning Environment
- Environment:Vespa engine Vespa Java 17 Build Runtime
- Environment:Marker Inc Korea AutoRAG API Keys Configuration
- Environment:SeldonIO Seldon core GPU Inference Environment
- Environment:OpenBMB UltraFeedback HuggingFace Hub Environment
- Environment:Volcengine Verl Ray Distributed Environment
- Environment:Openai Whisper PyTorch CUDA
- Environment:OpenGVLab InternVL Flash Attention 2
- Environment:Sktime Pytorch forecasting Cpflows MQF2 Dependencies