Home / Companies / LabelBox / Blog / Post Details
Content Deep Dive

Code Runner: Secure, scalable code execution for model evaluation

Blog post from LabelBox

Post Details
Company
Date Published
Author
Dmytro Apollonin
Word Count
866
Company Posts That Month
5
Language
-
Hacker News Points
-
Post removed?
No
Summary

Labelbox has introduced Code Runner, a new feature on its platform designed to enhance the evaluation of large language models (LLMs) by allowing users to execute code directly within the evaluation workflow. Code Runner aims to improve the quality of responses in coding-related projects by providing precise outputs, such as standard output, standard error, execution time, and warnings, without users needing to leave the platform. The infrastructure behind Code Runner is powered by Google Cloud Run, which offers a secure, scalable environment for executing code in isolated, temporary containers tailored to specific programming languages like Python and JavaScript. The system ensures security through measures such as separate Google Cloud Platform projects and communication via Private Service Connect, which prevents public exposure and restricts network access. Code Runner's architecture is designed for scalability, handling multiple requests efficiently, and reliability, as each execution occurs in a clean, stateless environment. By integrating this feature, Labelbox empowers users to perform dynamic, interactive testing, encouraging feedback and continuous improvement.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
AI Guardrails 2 186 50 28 +2%
LLM 2 2,668 436 137 -7%
Serverless 1 778 155 73 +74%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.