Search papers, labs, and topics across Lattice.
3
0
6
0
The single mid-run human intervention that pulled the community out of a monoculture is described, what the trace does and does not establish, and the controlled comparison that would settle whether shared research state improves discovery per unit of compute is described.
StepAudio 3 Realtime, an audio-language foundation model organized around a continuous listen-converse-think-act loop, achieves dialogue and reasoning performance comparable to dedicated reasoning models while speaking in real time, and resolves the tension between deep deliberation and latency via Think-While-Speaking.
This study introduces StepAudio 3 Gen, a general-purpose audio generation model that supports zero-shot text-to-speech (TTS), voice design, vocal generation, sound effects, music, vibe speech, and mixtures of multiple audio types within a unified framework.