Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:ClickHouse ClickHouse Running Stateless Tests
- Workflow:Isaac sim IsaacGymEnvs Custom Task Development
- Workflow:Sail sg LongSpec GLIDE Draft Model Training
- Workflow:CrewAIInc CrewAI Sequential Crew Execution
- Workflow:Apache Airflow Core Release Process
- Workflow:DistrictDataLabs Yellowbrick Feature Analysis and Selection
- Workflow:Datahub project Datahub Docker Quickstart Deployment
- Workflow:Rapidsai Cuml Sklearn Zero Code Acceleration
- Workflow:Allenai Open instruct Tulu3 Full Post Training
- Workflow:Googleapis Python genai Text Content Generation
Principles
- Principle:Huggingface Datasets Data Download and Preparation
- Principle:Langchain ai Langgraph Runtime Context
- Principle:NVIDIA NeMo Aligner Reward Model Serving
- Principle:Triton inference server Server Performance Analysis
- Principle:Google deepmind Dm control Task Registry Discovery
- Principle:PacktPublishing LLM Engineers Handbook Evaluation Results Aggregation
- Principle:DataTalksClub Data engineering zoomcamp Streaming Data Model
- Principle:Tensorflow Tfjs Model Inference
- Principle:ClickHouse ClickHouse Server Package Configuration
- Principle:Protectai Llm guard Secret Detection
Implementations
- Implementation:CARLA simulator Carla WorldSnapshot Class
- Implementation:Run llama Llama index EvaluatorEvaluationDataset
- Implementation:OpenRLHF OpenRLHF PairWiseLoss
- Implementation:Scikit learn Scikit learn Fetch20Newsgroups
- Implementation:Risingwavelabs Risingwave Docker Environment Config
- Implementation:FMInference FlexLLMGen DeepSpeed AIO Perf Sweep
- Implementation:Scikit learn Scikit learn PassiveAggressiveClassifier
- Implementation:Apache Airflow Configuration Parser
- Implementation:LMCache LMCache ZMQ Offload Server
- Implementation:Treeverse LakeFS Java SDK Model GarbageCollectionPrepareResponse
Heuristics
- Heuristic:Princeton nlp SimPO Multi Seed Diversity
- Heuristic:Avdvg InjectGuard Module Level Initialization
- Heuristic:MarketSquare Robotframework browser MacOS Sonoma Startup Delay
- Heuristic:Langchain ai Langchain Retry Scope Best Practice
- Heuristic:LMCache LMCache Prefix Based Retrieval Pattern
- Heuristic:ThreeSR Awesome Inference Time Scaling API Rate Limiting Tip
- Heuristic:Openai Whisper Temperature Fallback Strategy
- Heuristic:NVIDIA NeMo Aligner PPO NCCL Algorithm Setting
- Heuristic:LaurentMazare Tch rs Safetensors Format Preference
- Heuristic:Bitsandbytes foundation Bitsandbytes Compute Dtype Mismatch Warning
Environments
- Environment:Evidentlyai Evidently LLM Evaluation Environment
- Environment:Webdriverio Webdriverio Browser Driver Environment
- Environment:Duckdb Duckdb Release Publishing Env
- Environment:OpenGVLab InternVL PyTorch CUDA
- Environment:Spotify Luigi Apache Spark
- Environment:Marker Inc Korea AutoRAG VLLM Environment
- Environment:Heibaiying BigData Notes Spark 2 4 Environment
- Environment:Haifengl Smile Quarkus Serve Environment
- Environment:CrewAIInc CrewAI Python Runtime Environment
- Environment:Protectai Llm guard Python Runtime Dependencies