Reddit is testing a new video and audio experience that lets users listen to its most viral posts, drawing inspiration from short-form video platforms. At the same time, Smallest.ai raised $13M to build voice AI models explicitly designed to make AI phone calls pass the Turing test. These two pivots share an axis: the interface is moving from text-read to voice-heard, and the question of whether what you're hearing is human is becoming functionally irrelevant to the experience.
Why Audio Became the New Scroll
The logic is not mysterious. Screens require attention. Audio requires presence. As commuting revived and ambient media consumption normalized post-pandemic, the ear became the underserved organ of the feed economy. Reddit's move mirrors a broader platform calculus: if TikTok proved that the brain processes short video differently than text, the next bet is that audio-first content captures idle attention that neither text posts nor video can reach. The Reuters Institute's 2026 Digital News Report, cited by Fast Company, found declining trust in AI-generated answers. Yet the same report documents rising audio news consumption. The ear trusts in a register the eye has become skeptical of.
When the Human-Sounding AI Calls You Back
Smallest.ai's pitch is explicit: build voice models that pass the Turing test in real phone interactions. The product category (AI voice for customer service and outbound calls) already exists, but the framing around passing for human marks a shift from utility to mimicry. A 2026 paper by Fenoglio in arXiv CS.CY argues that contemporary AI discourse routinely attributes to language models properties they cannot bear, including genuine understanding. Yet the commercial imperative runs opposite: make it sound like it can bear them. Reddit's audio feed and Smallest.ai's human-passing voice are not the same product, but they converge on the same cultural logic. Soleio's observation that speed is a moat and the end of the human monopoly on taste is near applies just as cleanly to voice: once AI sound becomes indistinguishable, human-sounding becomes a style choice, not a guarantee.