Light
Home
/
Companies
/
AssemblyAI
/
Hacker News
AssemblyAI on HN
61 posts with 1+ points since 2022
Filters
Min points:
1
10
25
50
100
250
500
Since:
2018
2019
2020
2021
2022
2023
2024
2025
2026
Posts by Month (61 total)
Hacker News Posts
Search:
Title
Points
Comments
Date
How DALL-E 2 Works
252
178
2022-04-19
How Imagen Works
142
97
2022-06-23
LeMUR: LLMs for Audio and Speech
129
23
2023-07-27
Build Your Own Imagen Text-to-Image Model
111
43
2022-08-17
The Full Story of Large Language Models and RLHF
108
20
2023-05-03
Introduction to Diffusion Models for Machine Learning
98
10
2022-05-12
How RLHF Preference Model Tuning Works (and How Things May Go Wrong)
95
9
2023-08-09
Using JAX in 2022
66
30
2022-02-15
How physics advanced generative AI
49
8
2023-04-19
What is residual vector quantization?
45
8
2023-09-04
A Beginner's Guide to TorchStudio, the PyTorch IDE
16
0
2022-03-28
MediaPipe for Dummies
11
0
2022-05-02
Universal-1: Robust and accurate multilingual speech-to-text
11
0
2024-04-03
Differentiable Programming – A Simple Introduction
8
0
2022-03-02
Decoding Strategies – How LLMs Choose the Next Word
8
0
2024-08-21
Variational Autoencoders for Dummies
7
0
2022-01-03
How to Run Stable Diffusion to Generate Images (Local and Cloud)
7
0
2022-08-23
Emergent Abilities of Large Language Models
7
0
2023-03-07
Kaldi Speech Recognition – A Simple Tutorial
6
0
2022-01-20
Image generation – PFGM vs. Diffusion models?
6
0
2022-11-01
AlphaTensor explained – Motivation, method, and assessment
5
3
2022-11-22
How to evaluate Speech Recognition AI models
5
0
2023-06-16
Transcribe audio to text with Cloudflare Workers and AssemblyAI
5
0
2023-08-03
Conformer-2 AI model for speech recognition
5
0
2023-07-20
AI Generative Models for Audio
5
1
2023-07-16
How ChatGPT actually works
4
1
2023-04-27
What AI Music Generators Can Do (and How They Do It)
4
0
2023-09-22
Stable Diffusion faster in Keras thanks to XLA
4
0
2022-11-30
Automatically determine video sections with AI using Python
4
0
2023-11-07
Universal-3 Pro Streaming
3
0
2026-03-04
Complete guide to modern generative AI image models
3
None
2023-05-10
Introduction to Generative AI
3
None
2023-05-02
AI trends in 2023: Graph Neural Networks
3
None
2023-03-29
Conformer-1: a robust speech recognition model
3
None
2023-03-15
Why You Should (or Shouldn't) be Using Google's JAX in 2023
3
None
2023-02-11
Stable Diffusion 1 vs. 2 – What you need to know
3
None
2022-12-06
OpenAI Whisper Benchmarks
3
None
2022-09-27
MinImagen – Build Your Own Imagen Text-to-Image Model
3
None
2022-09-08
Variational Autoencoders Simply Explained
3
None
2022-04-06
We Built a Scalable AI Lakehouse at AssemblyAI
3
1
2024-11-22
Golden Gemini – A new approach in Speech AI
3
0
2025-02-04
How to Build an Audio Intelligence Dashboard
2
None
2022-09-21
RLHF vs. RLAIF for language model alignment
2
None
2023-08-22
AssemblyAI (YC S17) raises $50M Series C to build superhuman Speech AI …
2
None
2023-12-05
Model Context Protocol (MCP) – What it is, how it works, and …
2
None
2025-04-22
Five Things to know about Large Language Models
2
None
2023-05-23
AssemblyAI launches Voice Agent API – an end-to-end voice agent pipeline
2
None
2026-04-30
Universal-3.5 Pro Realtime
2
None
2026-06-23
AssemblyAI's (YC S17) Medical Mode: 20% fewer missed entities on medical terms
2
None
2026-03-25
Reinforcement Learning from AI Feedback
2
None
2023-08-01
Unsupervised Machine Learning for Beginners
2
None
2023-02-19
Introduction to LLMs for Generative AI
2
None
2023-05-17
AssemblyAI announces lower latency, lower cost, more possibilities
1
None
2024-01-11
AI for Universal Audio Understanding: Qwen-Audio Explained
1
None
2023-12-14
Retrieval Augmented Generation (Rag) on Audio Data with LangChain
1
None
2023-09-26
Free Speech-to-Text APIs, AI Models, and Open Source Engines
1
None
2022-12-11
Getting Started with Hugging Face's Gradio
1
None
2022-10-05
JavaScript Text-to-Speech – The Easy Way
1
None
2022-04-05
How Microsoft's New Large Vision Model "Florence-2" Works
1
None
2024-07-15
Universal-Streaming – built for AI voice agents
1
None
2025-06-11
Promptable Speech Language Model by AssemblyAI (YC S17)
1
None
2026-02-04