Skip to main content

The news desk

AI news, checked against the source.

What changed in AI, one story at a time. Each one links to its primary source, and facts stay apart from interpretation.

Last updated

Friday, October 2UTC

  1. New
    OpenAIOpenAI Python SDK 3.24.0Audio

    OpenAI Python SDK adds custom-voice creation and agent-session events

    The official release adds custom-voice creation support and session-event types, plus request-routing repairs. Custom voices remain restricted to eligible organizations and require consent and matching sample recordings. The SDK change is not unrestricted voice cloning or a new universal voice-access rollout.

    OpenAI3 sources

Free, email only

Unlock the full brief free.

  • Thirty days of stories to browse, not seven
  • On every story: what to check, each source and why it matters
  • The same signup unlocks every free tool

This is not an account: there is no password, and the unlock is a cookie in this browser. You also join the Rise Productive newsletter from Demetri Panici, about once a week: what I built and what changed in AI. We'll email you a link to confirm, and you can unsubscribe in one click. The same signup unlocks every free tool on the site. How your email is handled.

Already subscribed? Enter the same email to unlock this browser. You won't be signed up twice.

Thursday, October 1UTC

  1. New
    SunoSuno Speech betaAudio

    Suno Speech beta generates spoken audio and background music together

    Suno announced Speech, a beta model that combines spoken voice and original background music into one track. Users enter an idea or written text and describe a voice and musical style. After a limited test, the publisher says the beta is opening to everyone. It explicitly warns of accent drift and exaggerated pauses. The announcement does not establish a speech API, voice-cloning feature or commercial-use entitlement.

    Suno2 sources

  2. New
    MicrosoftMAI-Transcribe-2-Streaming and MAI-Voice-2.1Audio

    Microsoft pairs streaming transcription with multilingual MAI voices for conversational agents

    Microsoft AI launched MAI-Transcribe-2-Streaming alongside MAI-Voice-2.1 and its Flash variant. Streaming recognition supports 60 languages with continuously revised partial transcripts; the voice models support 23 languages and 26 locales, with a consistent speaker identity across languages. Microsoft lists Foundry, Playground, Vercel and Azure Voice Live access, with LiveKit still coming soon. Personal voice cloning requires gated approval and recorded talent consent. These are audio components for an agent loop, not an autonomous agent or independently validated speed record.

    Microsoft AI4 sources

Days are UTC. Some sources report an exact time and others only a date; both are placed on their UTC day.