Updated monthly

Changelog

New models, features, and API changes shipped to Venice each month.

Venice

Changelog

May 2026

Models

47 changes

Text models14

  • DeepSeek V4 Flash

    Lighter, faster variant of DeepSeek V4 optimized for speed and lower latency while retaining strong general-purpose performance. Available to all users.

    All users
  • DeepSeek V4 Pro

    DeepSeek's full-size V4 reasoning model with extended context and strong performance on coding, math, and multi-step tasks. Available to all users.

    All users
  • Gemini 3.5 Flash

    Google DeepMind's lightweight, low-latency text model optimized for speed. Available to all users.

    All usersAnonymous
  • Gemma 4 26B A4B Uncensored

    Uncensored, unfiltered mixture-of-experts variant of Google's Gemma 4 with 26B total parameters and 4B active parameters. Available to all users.

    All usersTEE
  • Gemma 4 31B Instruct

    Google's 31B-parameter open text model with instruction tuning. Available to all users.

    All users
  • GLM 5.1 E2EE

    Zhipu AI's GLM 5.1 running with end-to-end encryption in a Trusted Execution Environment. Available to Pro users at no additional credit cost.

    Pro users

Image models6

  • Grok Imagine High Quality

    Image generation model from xAI with state-of-the-art quality output. Available to all users.

    All usersPrivate
  • Wan 2.7 Pro Edit

    Alibaba DashScope image editing model for prompt-driven edits to existing images. Available to all users.

    All users
  • GPT Image 2

    New quality setting added for GPT Image 2 image generation.

    All users
  • GPT Image 2 Quality Parameter

    GPT Image 2 now supports a quality parameter with pricing tied to resolution and quality.

    All users
  • Lustify v7

    Restored Lustify v7 model availability after prior deprecation.

    All users
  • Qwen Image

    Effective June 18, Qwen Image pricing will increase from $0.01 to $0.03 per generated image. The model will also move from height / width parameters to aspect_ratio, with support for: 1:1, 3:2, 16:9, 21:9, 9:16, 2:3, 3:4, and 4:5.

    All users

Video models12

  • HappyHorse 1.0

    Alibaba's text-to-video generation model. Available to all users.

    All users
  • HappyHorse 1.0 Edit

    Video editing model for modifying and transforming existing video. Available to all users.

    All users
  • HappyHorse 1.0 I2V

    Image-to-video generation model from a source image. Available to all users.

    All users
  • HappyHorse 1.0 Reference

    Video generation guided by a reference image for style and content. Available to all users.

    All users
  • Kling O3 4K

    Kuaishou O3-series text-to-video model generating native 4K resolution output. Available to all users.

    All users
  • Kling O3 4K I2V

    Kuaishou O3-series image-to-video model generating native 4K resolution output. Available to all users.

    All users

Audio models3

  • Chatterbox HD on /models

    Chatterbox HD voice cloning model is now listed and documented on the /models endpoint.

    All users
  • Lyria 3 Pro

    Google DeepMind audio and music generation model capable of producing high-fidelity instrumental and vocal tracks. Available to all users.

    All usersAnonymous
  • Music Pricing Granularity

    Character-based music pricing now uses finer 100-character granularity.

    All users

Deprecated12

  • GLM 5
  • Grok 4.1 Fast
  • HiDream
  • Kimi K2 Thinking
  • Kimi K2 Thinking
  • Qwen 3 Coder 480B
  • Qwen 3.5 122B A10B
  • Qwen3 Coder 480B
  • NEAR AI GLM 5.0 (E2EE)
  • Venice Uncensored 1.1
  • Grok Imagine Pro
  • Qwen Image Deprecation

Features

112 changes

Chat31

  • Agentic Chat

    Agentic Chat is now the default Venice chat experience, with tool use, media generation, and multi-step workflows available directly inside a conversation. Users can search, reason, generate or edit images, create videos, and refine results without switching chats.

  • Agentic Chat Message Editing

    Messages in agentic chat can now be edited in place without resending.

  • Audio Reference Chips

    Audio files can now be attached as reference chips via the chat slash menu.

  • Auto-Approve Video in Agentic Chat

    Video generation requests in agentic chat are now auto-approved without requiring manual confirmation.

  • China Server Location Flag

    China flag icon now displayed for CN server locations in model details.

  • Drag-to-Folder for Agentic Chat

    Agentic chat conversations can now be dragged into folders in the sidebar.

Image, video and audio15

  • Download All Videos

    New 'Download All Videos' button in the Studio gallery to batch-download all generated videos.

  • Image Auto-Downsize on Share

    Images larger than 25 MB are automatically downsized before sharing.

  • Image Generation Metadata

    Generated images now include generation metadata (model, prompt, settings) embedded in the file.

  • Image-to-Video Progress Preview

    A blurred version of the source image is now displayed while image-to-video generation is in progress.

  • In-Browser Camera Capture

    Capture photos directly from the browser camera in Chat, Video Studio, Image Studio, and multi-edit.

  • Suggested Prompts for Image & Video

    Pre-written prompt suggestions now appear in the image and video generation interfaces.

Account and settings6

  • Conversation Mode Minutes Display

    Remaining conversation mode minutes are now shown in the Storage & Limits settings page.

  • Email Identity Verification

    Optional email address can now be linked for account identity verification.

  • Free User Rate Limit CTA

    Free users now see a call-to-action prompt when they hit rate limits.

  • Two-Factor Authentication

    Added additional second-factor authentication options for account security.

  • Pro Upgrade CTA

    Updated copy on the Pro upgrade call-to-action.

  • User Dropdown Menu

    Improved reordered items in the user dropdown menu.

Wallet and payments20

  • Auto Top-Up

    New option to automatically replenish credits when balance falls below a set threshold.

  • Credit Cost Granularity

    Credit costs now displayed with per-100-character precision.

  • Credit Usage History

    A detailed log of past credit usage is now available in the wallet section.

  • Crypto Subscription Credit Display

    Credit balance now shown on the crypto subscription card.

  • Crypto Subscription Management

    Users can now upgrade, downgrade, or cancel crypto subscriptions directly from the crypto subscription card in the wallet section.

  • Solana Top-Up Support

    Solana is now accepted as a payment method for credit top-ups and wallet authentication.

General9

  • Clickable Creator Name

    Creator name in the Social Feed post detail header now links to the creator's profile.

  • Emoji Reactions

    Emoji reactions are now available on social posts in the Social Feed.

  • Home Page Redesign

    Redesigned home page with updated layout at venice.ai/home.

  • Venice Skills GitHub Repository

    Official veniceai/skills repository now live on GitHub with example skills covering the full Venice API surface.

  • Wide Screen Layout

    Improved 2-column grid layout on wide screens for better use of available space.

  • Execution Time Display

    Updated execution time display to show milliseconds for greater precision.

Mobile app31

  • Agentic Chat

    Agentic Chat (Chat V2) is a new multi-step chat mode powered by an agent with access to Venice tools and features.

  • Background Chat Sync

    Chat responses that streamed while the app was backgrounded sync upon returning to the foreground.

  • Context Search Rendering

    Context search results are now rendered inline in chat.

  • Enter Key to Send

    Pressing Enter now submits messages in chat.

  • Android Native Chat Streaming

    Chat responses now stream using native Android processing.

  • iOS Native Chat Streaming

    Chat responses now stream using native iOS processing.

API

41 changes
  • OpenAI-Compatible File Inputs

    Chat completions endpoint now accepts file inputs using the OpenAI-compatible format.

  • maxtokens Strict Cap on Reasoning Models

    On reasoning-capable models, maxtokens is now a strict cap on total completion tokens (visible output + reasoning), restoring Venice's prior behavior across the model fleet. maxcompletiontokens is accepted as an equivalent alias and takes precedence if both are sent.

  • Image Edit Resolution Parameter

    New resolution parameter available on the image edit and multi-edit API endpoints.

  • Reference Audio URLs

    OpenAPI docs now expose referenceaudiourls for supported audio-reference workflows.

  • Seedance R2V Audio

    Seedance reference-to-video workflows can now use reference audio through the public API.

  • TTS Response Format Selection

    The text-to-speech endpoint now accepts a response format parameter to specify the output audio format per request.

Fixes

29 changes
  • Android Chat Reliability

    Fixed chat dropping or failing during request timeouts and mid-stream disconnects on Android.

    Mobile
  • Attachment-Only Messages

    Fixed inability to send messages containing only an attachment without text.

    Chat
  • Chat Message Queue

    Fixed chat message queue issues that could cause messages to be processed incorrectly.

    Chat
  • Conversation Replay Fix

    Fixed a bug where already-read responses would replay when re-entering a conversation.

    Mobile
  • Grok 4.1 Fast Characters

    Fixed errors when using Grok 4.1 Fast with characters.

    Chat
  • Quoting Video Content

    Fixed an error occurring when quoting video content in conversations.

    Chat

Have an idea for Venice?

Share feature requests and vote on ideas from the community on Learn Venice.

Submit an idea
Building on Venice?Read the API docsBrowse FAQs