Home / Companies / AssemblyAI / Hacker News

AssemblyAI on HN

61 posts with 1+ points since 2022

Filters
Since:
Posts by Month (61 total)
Hacker News Posts
Title Points Comments Date
How DALL-E 2 Works 252 178 2022-04-19
How Imagen Works 142 97 2022-06-23
LeMUR: LLMs for Audio and Speech 129 23 2023-07-27
Build Your Own Imagen Text-to-Image Model 111 43 2022-08-17
The Full Story of Large Language Models and RLHF 108 20 2023-05-03
Introduction to Diffusion Models for Machine Learning 98 10 2022-05-12
How RLHF Preference Model Tuning Works (and How Things May Go Wrong) 95 9 2023-08-09
Using JAX in 2022 66 30 2022-02-15
How physics advanced generative AI 49 8 2023-04-19
What is residual vector quantization? 45 8 2023-09-04
A Beginner's Guide to TorchStudio, the PyTorch IDE 16 0 2022-03-28
MediaPipe for Dummies 11 0 2022-05-02
Universal-1: Robust and accurate multilingual speech-to-text 11 0 2024-04-03
Differentiable Programming – A Simple Introduction 8 0 2022-03-02
Decoding Strategies – How LLMs Choose the Next Word 8 0 2024-08-21
Variational Autoencoders for Dummies 7 0 2022-01-03
How to Run Stable Diffusion to Generate Images (Local and Cloud) 7 0 2022-08-23
Emergent Abilities of Large Language Models 7 0 2023-03-07
Kaldi Speech Recognition – A Simple Tutorial 6 0 2022-01-20
Image generation – PFGM vs. Diffusion models? 6 0 2022-11-01
AlphaTensor explained – Motivation, method, and assessment 5 3 2022-11-22
How to evaluate Speech Recognition AI models 5 0 2023-06-16
Transcribe audio to text with Cloudflare Workers and AssemblyAI 5 0 2023-08-03
Conformer-2 AI model for speech recognition 5 0 2023-07-20
AI Generative Models for Audio 5 1 2023-07-16
How ChatGPT actually works 4 1 2023-04-27
What AI Music Generators Can Do (and How They Do It) 4 0 2023-09-22
Stable Diffusion faster in Keras thanks to XLA 4 0 2022-11-30
Automatically determine video sections with AI using Python 4 0 2023-11-07
Universal-3 Pro Streaming 3 0 2026-03-04
Complete guide to modern generative AI image models 3 None 2023-05-10
Introduction to Generative AI 3 None 2023-05-02
AI trends in 2023: Graph Neural Networks 3 None 2023-03-29
Conformer-1: a robust speech recognition model 3 None 2023-03-15
Why You Should (or Shouldn't) be Using Google's JAX in 2023 3 None 2023-02-11
Stable Diffusion 1 vs. 2 – What you need to know 3 None 2022-12-06
OpenAI Whisper Benchmarks 3 None 2022-09-27
MinImagen – Build Your Own Imagen Text-to-Image Model 3 None 2022-09-08
Variational Autoencoders Simply Explained 3 None 2022-04-06
We Built a Scalable AI Lakehouse at AssemblyAI 3 1 2024-11-22
Golden Gemini – A new approach in Speech AI 3 0 2025-02-04
How to Build an Audio Intelligence Dashboard 2 None 2022-09-21
RLHF vs. RLAIF for language model alignment 2 None 2023-08-22
AssemblyAI (YC S17) raises $50M Series C to build superhuman Speech AI … 2 None 2023-12-05
Model Context Protocol (MCP) – What it is, how it works, and … 2 None 2025-04-22
Five Things to know about Large Language Models 2 None 2023-05-23
AssemblyAI launches Voice Agent API – an end-to-end voice agent pipeline 2 None 2026-04-30
Universal-3.5 Pro Realtime 2 None 2026-06-23
AssemblyAI's (YC S17) Medical Mode: 20% fewer missed entities on medical terms 2 None 2026-03-25
Reinforcement Learning from AI Feedback 2 None 2023-08-01
Unsupervised Machine Learning for Beginners 2 None 2023-02-19
Introduction to LLMs for Generative AI 2 None 2023-05-17
AssemblyAI announces lower latency, lower cost, more possibilities 1 None 2024-01-11
AI for Universal Audio Understanding: Qwen-Audio Explained 1 None 2023-12-14
Retrieval Augmented Generation (Rag) on Audio Data with LangChain 1 None 2023-09-26
Free Speech-to-Text APIs, AI Models, and Open Source Engines 1 None 2022-12-11
Getting Started with Hugging Face's Gradio 1 None 2022-10-05
JavaScript Text-to-Speech – The Easy Way 1 None 2022-04-05
How Microsoft's New Large Vision Model "Florence-2" Works 1 None 2024-07-15
Universal-Streaming – built for AI voice agents 1 None 2025-06-11
Promptable Speech Language Model by AssemblyAI (YC S17) 1 None 2026-02-04