Home / Companies / Supabase / Blog / Post Details
Content Deep Dive

AI Inference now available in Supabase Edge Functions

Blog post from Supabase

Post Details
Company
Date Published
Author
Lakshan Perera
Word Count
997
Company Posts That Month
17
Language
English
Hacker News Points
33
Post removed?
No
Summary

We're making it super easy to run AI models within Supabase Edge Functions. A new built-in API is available within the Edge Runtime to run inference workloads in just a few lines of code, allowing developers to generate embeddings using models like `gte-small`, store and retrieve with pgvector, and use Large Language Models like `llama2` and `mistral`. This release solves technical challenges for developers who want to generate embeddings from their database content, offloading the compute-intensive task to background workers. Embedding generation uses the ONNX runtime under the hood, a cross-platform inferencing library that supports multiple execution providers from CPU to specialized GPUs. The new API is available today for all Edge Functions users in both local, hosted, and self-hosted platforms, with lower costs compared to proprietary LLMs like OpenAI and Claude, which charge per token or provide APIs to generate text embeddings. We're working with Ollama to make Large Language Models possible with Supabase: local development, self-hosted, and on the platform, with experimental support for `Llama` & `Mistral`.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Vector Search 26 2,722 279 102 +43%
LLM 6 3,669 412 154 +40%
Edge Computing 3 40 14 10 -18%
RAG 2 1,867 232 78 +54%
Developer Experience 1 287 186 98 -17%
Real-time 1 2,509 695 218 -9%
Serverless 1 1,024 191 85 +26%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.