---
title: "Improved Batch Inference API: Enhanced UI, Expanded Model Support, and 3000× Rate Limit Increase"
url: https://daily.dev/posts/improved-batch-inference-api-enhanced-ui-expanded-model-support-and-3000-rate-limit-increase-k9cc4jqs0
source_url: https://www.together.ai/blog/batch-inference-api-updates-2025
type: article
source: "Together AI"
published: 2026-05-31T07:40:44.592Z
updated: 2026-08-24T06:56:57.030Z
tags: ["llm"]
reading_time: 2
upvotes: 0
comments: 0
language: en
---

> ## Documentation Index
> Fetch the complete documentation index at: https://daily.dev/llms.txt
> Use this file to discover all available pages before exploring further.

# Improved Batch Inference API: Enhanced UI, Expanded Model Support, and 3000× Rate Limit Increase

**[Together AI](https://daily.dev/sources/togetherai)** · 2 min read · 0 upvotes · 0 comments

## Summary

Together.ai has rolled out major updates to its Batch Inference API, including a streamlined UI for creating and tracking batch jobs, universal support for all serverless models and private deployments, a 3000× rate limit increase (from 10M to 30B enqueued tokens per model per user), and pricing at 50% of the real-time API cost. Key use cases include large-scale text analysis, synthetic data generation, embedding generation, fraud detection, content moderation, model evaluation, and customer support automation.

## Full article

daily.dev links to this article rather than hosting it. Read it at the original source: <https://www.together.ai/blog/batch-inference-api-updates-2025>

## Similar posts on daily.dev

- [Scalable, Cost-Efficient AI: Introducing Unified Batch Inference on DigitalOcean](https://daily.dev/posts/scalable-cost-efficient-ai-introducing-unified-batch-inference-on-digitalocean-g2kea1ifv) · DigitalOcean · 1 upvotes · 0 comments
- [BigQuery enhancements to boost gen AI inference](https://daily.dev/posts/bigquery-enhancements-to-boost-gen-ai-inference-nddw0leg2) · Google Cloud · 1 upvotes · 0 comments

---

Tags: [#llm](https://daily.dev/tags/llm)

[View this post on daily.dev](https://daily.dev/posts/improved-batch-inference-api-enhanced-ui-expanded-model-support-and-3000-rate-limit-increase-k9cc4jqs0)
