Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Princeton nlp Tree of thought llm Adding new task
- Workflow:Onnx Onnx External Data Handling
- Workflow:Evidentlyai Evidently LLM Evaluation Monitoring
- Workflow:Apache Beam Twister2 Batch Execution
- Workflow:OWASP Www project top 10 for large language model applications Agentic Security Assessment
- Workflow:OpenRLHF OpenRLHF Iterative DPO
- Workflow:Allenai Open instruct Tulu3 Full Post Training
- Workflow:Apache Shardingsphere Shadow Rule Configuration
- Workflow:CrewAIInc CrewAI Sequential Crew Execution
- Workflow:Deepset ai Haystack Hybrid Document Search
Principles
- Principle:Princeton nlp Tree of thought llm Task Instantiation
- Principle:Online ml River TextClust Clustering
- Principle:FMInference FlexLLMGen DeepSpeed Package Build
- Principle:Mlfoundations Open flamingo Pretrained Weight Loading
- Principle:Cleanlab Cleanlab Issue Retrieval
- Principle:AUTOMATIC1111 Stable diffusion webui Embedding serialization
- Principle:Sktime Pytorch forecasting Synthetic Data Generation
- Principle:Mlc ai Web llm Structured Output Parsing
- Principle:OpenGVLab InternVL Length Grouped Sampling
- Principle:Hpcaitech ColossalAI Ray Weight Synchronization
Implementations
- Implementation:DataTalksClub Data engineering zoomcamp Java AvroProducer
- Implementation:Openai Openai python Resources Package Exports
- Implementation:Spotify Luigi BigQueryTarget
- Implementation:Microsoft Playwright AndroidDispatcher
- Implementation:LLMBook zh LLMBook zh github io Build Alibi Tensor
- Implementation:NVIDIA NeMo Aligner MegatronGPT Knowledge Distillation
- Implementation:Onnx Onnx Proto Utils
- Implementation:Mage ai Mage ai Chargebee Common Quotes Schema
- Implementation:Bitsandbytes foundation Bitsandbytes Linear4bit FSDP
- Implementation:NVIDIA TransformerEngine TELlamaForCausalLM
Heuristics
- Heuristic:Zai org CogVideo CPU Offload Strategy
- Heuristic:Liu00222 Open Prompt Injection PPL Threshold Tuning
- Heuristic:Run llama Llama index Batch Eval Retry Strategy
- Heuristic:Truera Trulens Temperature Zero For Deterministic Scoring
- Heuristic:ContextualAI HALOs TF32 Matmul Acceleration
- Heuristic:Microsoft BIPIA BF16 Compute Capability Check
- Heuristic:NVIDIA NeMo Aligner PPO NCCL Algorithm Setting
- Heuristic:Rapidsai Cuml GPU Cache Alignment
- Heuristic:Zai org CogVideo Frame Count and Resolution Constraints
- Heuristic:Volcengine Verl FSDP Mixed Precision Init
Environments
- Environment:CrewAIInc CrewAI Optional Provider Dependencies
- Environment:Protectai Llm guard Python Runtime Dependencies
- Environment:Astronomer Astronomer cosmos Cloud Provider Dependencies
- Environment:Deepspeedai DeepSpeed NVMe Environment
- Environment:Pytorch Serve CUDA GPU Environment
- Environment:Fede1024 Rust rdkafka Kafka Broker Runtime
- Environment:Hiyouga LLaMA Factory Optional Inference Backends
- Environment:OpenBMB UltraFeedback vLLM Multi GPU Environment
- Environment:Apache Airflow Development Contributor Environment
- Environment:Pyro ppl Pyro Python PyTorch Core