Fish Audio's anniversary post is the announcement behind this week's launch: a $52M seed led by Coreline Ventures and Capital Today, $21M ARR, more than 8 million users and over 2 million community voice models, built by a team that went from 3 people to 22 in a year. The part that changes your week is the offer attached to it. S2.1 Pro, the largest model, is free to developers via API through the end of August, with 50% off creator plans and three months free for migrations. The model claims 83 or more languages with native cadence, word-level emotion control across 15,000 natural language tags, more than 8,000 tokens per second on a single H200, and a 66% preference rate against competitors in blind listening tests. Note what the post does not do: it gives no per-character pricing to compare against ElevenLabs or Cartesia, only that inference optimisation dropped cost per request far enough to give the model away. Free until August 31 is a window to test in, not a cost you can plan a business on.
Rescript is a browser based video editor that lets you cut and rearrange footage by editing the transcript text, similar to Descript. It runs completely offline and locally, making it a privacy friendly option for creators. The tool is launching today on Product Hunt and available on GitHub.
Cseti's CrossView-Warp IC-LoRA for LTX-Video 2.3 takes an existing clip plus a numeric camera offset (azimuth, elevation, distance) and regenerates the scene from that viewpoint. A 3D orbit picker keyframes a full camera move across the clip, and the model is fed two references: a depth-warped copy that carries geometry, and the original that holds identity. The model card is honest about the catch. It steers the viewpoint rather than reprojecting geometry, so it routinely rotates less than you asked for. This is a v0.9 proof of concept trained on 548 real and synthetic camera pairs, and it needs the 22B LTX-Video 2.3 base under Lightricks' licence, so price the VRAM before you plan a shoot around a reshoot you cannot do.
A tutorial breaks down three methods for maintaining continuity across AI generated video scenes: continuing generation from a previous clip, frame chaining from the final image of a scene, and locking character details, outfits, and locations. The creator also demonstrates using Codex with the Higgsfield MCP to gather references and organize projects. The advice is actionable for daily production, though the video includes affiliate links to Higgsfield.