Changelog
Track all updates, improvements, and new features in InferenceIQ.
March 8, 2026
v2.4.0
Multi-Region Routing & New Providers
Major release bringing intelligent multi-region routing and support for 3 new LLM providers.
- FeatureAutomatic multi-region request routing with latency-aware load balancing
- FeatureSupport for Mistral AI, Cohere, and Replicate providers
- ImprovementReduced P99 latency by 28% through improved provider selection algorithm
- FixFixed edge case in fallback provider logic causing occasional timeout errors
February 28, 2026
v2.3.2
Latency Tracking & Bug Fixes
Enhanced monitoring and stability improvements.
- FeatureReal-time latency tracking dashboard for each provider
- ImprovementBetter error messages for authentication failures
- FixResolved issue with cache invalidation in concurrent requests
- FixFixed missing X-Request-ID headers in streaming responses
February 15, 2026
v2.3.0
Analytics Dashboard & Cost Forecasting
New analytics dashboard and AI-powered cost forecasting capabilities.
- FeatureComprehensive analytics dashboard with cost trends and provider comparison
- FeatureMachine learning-based cost forecasting for next 30-90 days
- FeatureCustom alert thresholds for cost overruns
- ImprovementFaster dashboard load times with data streaming
January 30, 2026
v2.2.1
Provider Failover Improvements
Enhanced reliability and failover mechanisms.
- ImprovementSmarter failover strategy with exponential backoff
- ImprovementBetter handling of rate-limited providers
- FixFixed stuck requests in failover scenario
January 15, 2026
v2.2.0
Python & TypeScript SDK v2.0
Major SDK release with TypeScript support and new streaming capabilities.
- FeatureOfficial TypeScript SDK with full type safety
- FeatureStreaming support in both Python and Node.js SDKs
- ImprovementPython SDK now uses async/await patterns
- BreakingDeprecated synchronous methods in Python SDK
December 20, 2025
v2.1.0
Real-Time Alerts & Slack Integration
Monitor your costs with real-time alerts and Slack notifications.
- FeatureReal-time cost alerts with configurable thresholds
- FeatureSlack integration for instant notifications
- FeatureEmail digests with daily/weekly summaries
- ImprovementBetter provider health status reporting
November 1, 2025
v2.0.0
Platform Rewrite & New API
Complete platform rewrite with new API architecture and improved performance.
- FeatureRedesigned REST API with better error handling
- FeatureSupport for 15+ new LLM providers
- Improvement50% reduction in API latency
- Breakingv1.x API endpoints are deprecated
October 15, 2025
v1.5.0
Initial Public Beta
InferenceIQ is now in public beta! Join us in optimizing AI inference costs.
- FeaturePublic beta launch with core routing functionality
- FeatureBasic dashboard and usage analytics
- FeatureSupport for OpenAI, Anthropic, and Google Cloud providers
- FeatureFree tier with 1M monthly inference requests