INSTRUCTIONS FOR LITELLM

This document provides comprehensive instructions for AI agents working in the LiteLLM repository.

OVERVIEW

LiteLLM is a unified interface for 100+ LLMs that:

Translates inputs to provider-specific completion, embedding, and image generation endpoints
Provides consistent OpenAI-format output across all providers
Includes retry/fallback logic across multiple deployments (Router)
Offers a proxy server (LLM Gateway) with budgets, rate limits, and authentication
Supports advanced features like function calling, streaming, caching, and observability

Provider Implementations: When adding/modifying LLM providers:
- Follow existing patterns in litellm/llms/{provider}/
- Implement proper transformation classes that inherit from BaseConfig
- Support both sync and async operations
- Handle streaming responses appropriately
- Include proper error handling with provider-specific exceptions
Type Safety:
- Use proper type hints throughout
- Update type definitions in litellm/types/
- Ensure compatibility with both Pydantic v1 and v2
Testing:
- Add tests in appropriate tests/ subdirectories
- Include both unit tests and integration tests
- Test provider-specific functionality thoroughly
- Consider adding load tests for performance-critical changes

Function/Tool Calling:
- LiteLLM standardizes tool calling across providers
- OpenAI format is the standard, with transformations for other providers
- See litellm/llms/anthropic/chat/transformation.py for complex tool handling
Streaming:
- All providers should support streaming where possible
- Use consistent chunk formatting across providers
- Handle both sync and async streaming
Error Handling:
- Use provider-specific exception classes
- Maintain consistent error formats across providers
- Include proper retry logic and fallback mechanisms
Configuration:
- Support both environment variables and programmatic configuration
- Use BaseConfig classes for provider configurations
- Allow dynamic parameter passing

The proxy server is a critical component that provides:

Key files:

LiteLLM supports MCP for agent workflows:

MCP server integration for tool calling
Transformation between OpenAI and MCP tool formats
Support for external MCP servers (Zapier, Jira, Linear, etc.)
See litellm/experimental_mcp_client/ and litellm/proxy/_experimental/mcp_server/