-
The University of Sheffield
- Sheffield
-
09:15
(UTC +01:00) - http://www.robertflynn.co.uk
- @RobFlynnHere
- https://huggingface.co/rjflynn2
- https://wandb.ai/wobrob101
Highlights
- Pro
Lists (1)
Sort Name ascending (A-Z)
Stars
The Full-Duplex Interaction Track of the ICASSP 2026 Human-like Spoken Dialogue Systems Challenge aims to advance the evaluation of full-duplex dialogue systems by in- troducing a dual-channel dial…
A library for preparing data for machine translation research (monolingual preprocessing, bitext mining, etc.) built by the FAIR NLLB team.
AI Audio Datasets (AI-ADS) 🎵, including Speech, Music, and Sound Effects, which can provide training data for Generative AI, AIGC, AI model training, intelligent audio tool development, and audio a…
A pool of speakers from Emilia-YODAS from open source data, suitable for multi-speaker voice applications.
A native-PyTorch library for large scale M-LLM (text/audio) training with tp/cp/dp.
Pipecat voice AI agents running locally on macOS
A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.
Un-0: an image generator powered by a simulated system of coupled oscillators, an example of an emerging physical computing substrate.
A TTS that fits in your CPU (and pocket)
Unified automatic quality assessment for speech, music, and sound.
Telegram ↔ tmux/herdr/agterm bridge for Claude Code, Codex CLI, and Pi Agent. Monitor output, respond to prompts, manage parallel sessions. Control AI coding agents from your phone.
The AI that really does things. Any OS. Any Platform. The lobster way. 🦞
Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work instead of supervising coding agents.
A lightweight alternative to OpenClaw that runs in containers for security. Connects to WhatsApp, Telegram, Slack, Discord, Gmail and other messaging apps,, has memory, scheduled jobs, and runs dir…
Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-agent workflows, automation, and chat apps
Connect your local Claude Code agent with Slack
Archived — ML Intern is no longer maintained. Continue with HuggingChat.
Kimi-Audio, an open-source audio foundation model excelling in audio understanding, generation, and conversation
SpeechJudge: Towards Human-Level Judgment for Speech Naturalness (https://arxiv.org/abs/2511.07931)
[NeurIPS' 25] Benchmark for evaluating TTS models on complex prosodic, expressiveness, and linguistic challenges.
[ACL 2024] Official PyTorch code for extracting features and training downstream models with emotion2vec: Self-Supervised Pre-Training for Speech Emotion Representation
Sync papers from Zotero to a reMarkable tablet
A curated list of projects related to the reMarkable tablet



