<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:wfw="http://wellformedweb.org/CommentAPI/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	xmlns:slash="http://purl.org/rss/1.0/modules/slash/"
	>

<channel>
	<title>TTS &#8211; Firethering</title>
	<atom:link href="https://firethering.com/tag/tts/feed/" rel="self" type="application/rss+xml" />
	<link>https://firethering.com</link>
	<description>Firethering is Your Hub for AI, Open Source and Tech That Actually Matters</description>
	<lastBuildDate>Thu, 14 May 2026 12:03:59 +0000</lastBuildDate>
	<language>en-US</language>
	<sy:updatePeriod>
	hourly	</sy:updatePeriod>
	<sy:updateFrequency>
	1	</sy:updateFrequency>
	<generator>https://wordpress.org/?v=6.6.9</generator>

<image>
	<url>https://firethering.com/wp-content/uploads/2024/10/cropped-firethering-FTR-favicon-32x32.png</url>
	<title>TTS &#8211; Firethering</title>
	<link>https://firethering.com</link>
	<width>32</width>
	<height>32</height>
</image> 
	<item>
		<title>DramaBox: An Open-Weight TTS Model Built Around Stage Directions</title>
		<link>https://firethering.com/dramabox-open-weights-tts-voice-cloning/</link>
					<comments>https://firethering.com/dramabox-open-weights-tts-voice-cloning/#respond</comments>
		
		<dc:creator><![CDATA[Mohit Geryani]]></dc:creator>
		<pubDate>Thu, 14 May 2026 11:55:37 +0000</pubDate>
				<category><![CDATA[Tech]]></category>
		<category><![CDATA[AI Models]]></category>
		<category><![CDATA[Trends]]></category>
		<category><![CDATA[AI]]></category>
		<category><![CDATA[TTS]]></category>
		<guid isPermaLink="false">https://firethering.com/?p=6810</guid>

					<description><![CDATA[Dramabox just landed on Hugging Face and the demo space is live. Resemble AI built it on top of Lightricks' LTX-2.3, and the thing that makes it different from every other TTS model is simpler than you'd expect, you don't give it text to read. You write it a scene.]]></description>
		
					<wfw:commentRss>https://firethering.com/dramabox-open-weights-tts-voice-cloning/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		<enclosure url="https://storage.googleapis.com/resemble-sampletables/Apr16/efqN7-b6HWE/ltx-tts-eval/expressive/generated/09_villain_sinister_laugh.wav" length="5285838" type="audio/wav" />
<enclosure url="https://storage.googleapis.com/resemble-sampletables/Apr24/mLbkPu2Qzwo/refs/002_ltx_tts_8ng372ra.wav" length="4461438" type="audio/wav" />
<enclosure url="https://storage.googleapis.com/resemble-sampletables/Apr24/mLbkPu2Qzwo/generated/002_ltx_tts_8ng372ra.wav" length="7251918" type="audio/wav" />

			</item>
		<item>
		<title>MOSS-TTS-Nano: Real-Time Voice AI on CPU, Part of an Open-Source Stack Rivaling Gemini</title>
		<link>https://firethering.com/moss-tts-nano-open-source-tts/</link>
					<comments>https://firethering.com/moss-tts-nano-open-source-tts/#respond</comments>
		
		<dc:creator><![CDATA[Mohit Geryani]]></dc:creator>
		<pubDate>Wed, 15 Apr 2026 08:51:04 +0000</pubDate>
				<category><![CDATA[Tech]]></category>
		<category><![CDATA[AI Models]]></category>
		<category><![CDATA[AI]]></category>
		<category><![CDATA[TTS]]></category>
		<guid isPermaLink="false">https://firethering.com/?p=6258</guid>

					<description><![CDATA[Most text-to-speech tools fall into two camps. The ones that sound good need serious hardware. The ones that run on anything sound robotic. MOSS-TTS-Nano is trying to be neither.

It's a 100 million parameter model that runs on a regular CPU and it actually sounds good. Good enough that the team behind it built an entire family of speech models around the same core technology, one of which has gone head to head with Gemini 2.5 Pro and ElevenLabs and come out ahead on speaker similarity.

It just dropped on April 10th and it's the newest addition to the MOSS-TTS family, a collection of five open source speech models from MOSI.AI and the OpenMOSS team. The family doesn't just cover lightweight local deployment. One of its models MOSS-TTSD outperforms Gemini 2.5 Pro and ElevenLabs on speaker similarity in benchmarks. Another generates voices purely from text descriptions with no reference audio needed. And one is built specifically for real-time voice agents with a 180ms first-byte latency.

Nano is the entry point. The family is the story.]]></description>
		
					<wfw:commentRss>https://firethering.com/moss-tts-nano-open-source-tts/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		<enclosure url="https://openmoss.github.io/MOSS-TTS-Nano-Demo/assets/%F0%9F%87%BA%F0%9F%87%B8%20A%20Gentle%20Reminder.wav" length="6051918" type="audio/wav" />

			</item>
		<item>
		<title>4 Open-Source TTS Models That Can Clone Voices and Actually Sound Human</title>
		<link>https://firethering.com/open-source-tts-voice-cloning/</link>
					<comments>https://firethering.com/open-source-tts-voice-cloning/#respond</comments>
		
		<dc:creator><![CDATA[Mohit Geryani]]></dc:creator>
		<pubDate>Sun, 05 Apr 2026 14:30:17 +0000</pubDate>
				<category><![CDATA[AI Picks]]></category>
		<category><![CDATA[Tech]]></category>
		<category><![CDATA[AI]]></category>
		<category><![CDATA[AI Models]]></category>
		<category><![CDATA[TTS]]></category>
		<guid isPermaLink="false">https://firethering.com/?p=6057</guid>

					<description><![CDATA[Voice cloning used to mean expensive studio software, proprietary APIs with per-character pricing, or models so heavy they needed server infrastructure just to run. That changed quietly over the last few months.

Four open source models exist right now that do something the previous generation struggled with. They do not just generate speech. They clone a voice from a short audio sample and produce output that is genuinely difficult to compare from the original speaker. 

The gap between open source and commercial TTS has been closing for a while. These four models suggest it has effectively closed for voice cloning specifically. Here is what each one actually does and who it is for.]]></description>
		
					<wfw:commentRss>https://firethering.com/open-source-tts-voice-cloning/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		<enclosure url="https://zhu-han.github.io/omnivoice/audios/seedtts/generated/seedtts_ref_en_3_gen_en.wav" length="629838" type="audio/wav" />
<enclosure url="https://aria-k-alethia.github.io/LongCat-AudioDiT-demo/audio/main/audio/5.wav" length="335916" type="audio/wav" />
<enclosure url="https://fireredteam.github.io/demos/firered_tts_2/audios/zero-shot-podcast-generation_samples/en_1_fireredtts.wav" length="3156524" type="audio/wav" />
<enclosure url="https://raw.githubusercontent.com/NextGenToolbox/AIdemvids/refs/heads/main/fishaudios2demo.mp3" length="118646" type="audio/mpeg" />

			</item>
	</channel>
</rss>
