Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Isaac sim IsaacGymEnvs Factory Assembly Training
- Workflow:Deepspeedai DeepSpeed Sequence Parallel Long Context Training
- Workflow:DistrictDataLabs Yellowbrick Feature Analysis and Selection
- Workflow:Dagster io Dagster Bluesky Analytics
- Workflow:OWASP Www project top 10 for large language model applications LLM Vulnerability Assessment
- Workflow:Duckdb Duckdb Extension Development And Distribution
- Workflow:Recommenders team Recommenders Neural Collaborative Filtering
- Workflow:Deepset ai Haystack RAG Evaluation Pipeline
- Workflow:SeleniumHQ Selenium Page Object Pattern Testing
- Workflow:PrefectHQ Prefect Dbt Model Orchestration
Principles
- Principle:Princeton nlp SimPO Chat Template Application
- Principle:Huggingface Trl Reward Preference Dataset Loading
- Principle:Cohere ai Cohere python Tool Result Submission
- Principle:Openclaw Openclaw Extension Type Identification
- Principle:Microsoft LoRA HF Trainer LoRA Training
- Principle:Scikit learn Scikit learn Score Distribution Analysis
- Principle:ThreeSR Awesome Inference Time Scaling Maintainer Review
- Principle:Apache Dolphinscheduler Failover Process Initiation
- Principle:Huggingface Datasets Parquet Import
- Principle:Bigscience workshop Petals Block Selection
Implementations
- Implementation:OWASP Www project top 10 for large language model applications VulnerabilityEntry Extract Attack Scenarios
- Implementation:Dagster io Dagster Dbt Asset Check Integration
- Implementation:Vibrantlabsai Ragas AnswerCorrectness
- Implementation:Cypress io Cypress AddTestingTypeToCypressConfig
- Implementation:Hiyouga LLaMA Factory Multimodal Plugin
- Implementation:Mistralai Client python FineTuningJobs Get List
- Implementation:Lance format Lance Tags CRUD
- Implementation:SeleniumHQ Selenium Architecture
- Implementation:Trailofbits Fickling FicklingMLUnpickler Load
- Implementation:CARLA simulator Carla Python TrafficManager Bindings
Heuristics
- Heuristic:FMInference FlexLLMGen Offloading Percent Tuning
- Heuristic:Ollama Ollama Multimodal Parallel Restriction
- Heuristic:Microsoft LoRA Selective LoRA QV Only
- Heuristic:Shiyu coder Kronos Learning Rate And Optimizer Tuning
- Heuristic:Fede1024 Rust rdkafka Queue Buffering Priority
- Heuristic:Mlc ai Mlc llm Engine Mode Selection
- Heuristic:Huggingface Trl Disable Dropout For RL Training
- Heuristic:Anthropics Anthropic sdk python Streaming For Long Requests
- Heuristic:Snorkel team Snorkel NLP Preprocessor Memoization
- Heuristic:Huggingface Trl DeepSpeed ZeRO3 Generation Tradeoff
Environments
- Environment:Dotnet Machinelearning Dotnet SDK And Runtime
- Environment:Marker Inc Korea AutoRAG Japanese NLP Dependencies
- Environment:Iterative Dvc DVC Environment Variables
- Environment:Mlc ai Mlc llm CUDA GPU Environment
- Environment:Mlflow Mlflow MLflow Server Environment
- Environment:Sgl project Sglang Grafana
- Environment:TobikoData Sqlmesh Dbt Compatibility
- Environment:Risingwavelabs Risingwave Python Tooling Environment
- Environment:Openai Evals OpenAI API Configuration
- Environment:Datajuicer Data juicer LLM API Credentials Environment