Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Dagster io Dagster LLM Fine Tuning
- Workflow:Facebookresearch Habitat lab Custom Task Extension
- Workflow:Interpretml Interpret EBM Model Merging
- Workflow:NVIDIA TransformerEngine Accelerate HF Llama With TE
- Workflow:Huggingface Peft Seq2Seq AdaLoRA Finetuning
- Workflow:LLMBook zh LLMBook zh github io LLM Pretraining
- Workflow:Pola rs Polars Data IO and Format Conversion
- Workflow:CarperAI Trlx RLHF Summarization Pipeline
- Workflow:Pola rs Polars DataFrame Aggregation and Grouping
- Workflow:Triton inference server Server Model Performance Tuning
Principles
- Principle:Apache Shardingsphere Configuration Object Conversion
- Principle:Lm sys FastChat FSDP Safe Model Saving
- Principle:Deepspeedai DeepSpeed Evoformer Attention Kernels
- Principle:Bentoml BentoML Model Cloud Sync
- Principle:Teamcapybara Capybara Form Field Filling
- Principle:Astronomer Astronomer cosmos Profile Configuration
- Principle:Cleanlab Cleanlab Synthetic Noise Generation
- Principle:Huggingface Datasets AudioFolder Dataset Building
- Principle:Recommenders team Recommenders Stratified Data Splitting
- Principle:Kubeflow Kubeflow Monitor And Iterate
Implementations
- Implementation:Duckdb Duckdb VergeSort
- Implementation:Microsoft Semantic kernel Concepts OpenAPI Resource
- Implementation:Neuml Txtai TextToSpeech
- Implementation:Triton inference server Server Trtllm Build
- Implementation:Togethercomputer Together python Model Types
- Implementation:NVIDIA NeMo Aligner Preprocess HelpSteer2 Data
- Implementation:Apache Beam JobServicePipelineResult
- Implementation:Huggingface Optimum ExporterConfig Validation
- Implementation:Mlflow Mlflow Trace Decorator
- Implementation:Langgenius Dify Refactor Component
Heuristics
- Heuristic:Deepspeedai DeepSpeed FP16 Convergence Tips
- Heuristic:Microsoft Onnxruntime Convergence Debugging Tips
- Heuristic:Isaac sim IsaacGymEnvs DR Setup Only Flag
- Heuristic:PrefectHQ Prefect Task Timeout Thread Limitation
- Heuristic:Evidentlyai Evidently Drift Detection Thresholds
- Heuristic:Nautechsystems Nautilus trader Strategy On Start Initialization
- Heuristic:OpenRLHF OpenRLHF Off Policy IS Correction Tip
- Heuristic:Neuml Txtai Thread Safety Constraints
- Heuristic:Huggingface Datatrove MinHash Parameter Tuning
- Heuristic:CARLA simulator Carla Traffic Manager Sync Mode
Environments
- Environment:Vllm project Vllm GitHub
- Environment:Spotify Luigi Apache Spark
- Environment:Pola rs Polars GPU Execution Environment
- Environment:Apache Druid Druid Operator Kubernetes
- Environment:Recommenders team Recommenders GPU CUDA Environment
- Environment:Princeton nlp SimPO VLLM Inference
- Environment:Unstructured IO Unstructured Profiling Tools
- Environment:Huggingface Open r1 CUDA Environment
- Environment:Turboderp org Exllamav2 Build Toolchain
- Environment:Kubeflow Pipelines Kubernetes Cluster