Home / Companies / Baseten / Blog / Post Details
Content Deep Dive

Introducing function calling and structured output for open-source and fine-tuned LLMs

Blog post from Baseten

Post Details
Company
Date Published
Author
Bryce Dubayah, Philip Kiely
Word Count
604
Company Posts That Month
4
Language
English
Hacker News Points
-
Post removed?
No
Summary

A new feature has been introduced in TensorRT-LLM Engine Builder to generate structured output during LLM inference. This includes JSON mode, where model output matches a given JSON schema, and function calling, where the LLM selects from provided tools to accomplish a task. Both functionalities have no marginal impact on tokens per second and are available for all LLMs deployed using the Engine Builder. The new features aim to address challenges in integrating LLMs with structured data, enabling developers to call LLMs with guaranteed output structure while adding negligible latency.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
LLM 23 3,889 441 129 +7%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.