|
LLMs in Prod Comes to Bangalore
|
Vrushank Vyas |
2024-07-29 |
1,045 |
--
|
|
AI audit checklist for internal AI platforms & enablement teams
|
Drishti Shah |
2025-12-10 |
2,089 |
--
|
|
May: Major Updates
|
Vrushank Vyas |
2024-05-31 |
355 |
--
|
|
HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in HuggingFace - …
|
The Quill |
2023-04-20 |
197 |
--
|
|
Agent observability: measuring tools, plans, and outcomes
|
Drishti Shah |
2025-11-28 |
1,619 |
--
|
|
My Journey with AI-Driven Development: From Curiosity to Necessity
|
Ayush |
2023-08-18 |
602 |
--
|
|
We're Afraid Language Models Aren't Modeling Ambiguity - Summary
|
The Quill |
2023-05-07 |
225 |
--
|
|
Language Models are Few-Shot Learners - Summary
|
Rohit Agarwal |
2023-04-15 |
230 |
--
|
|
What It Means To Go To Prod
|
Rohit Agarwal |
2024-04-13 |
312 |
--
|
|
Instruction Tuning with GPT-4 - Summary
|
The Quill |
2023-04-16 |
241 |
--
|
|
CAMEL: Communicative Agents for "Mind" Exploration of LLMs - Summary
|
Rohit Agarwal |
2023-04-14 |
352 |
--
|
|
LLMs in Prod: The Reality of AI Outages, No LLM is Immune
|
Siddharth Sambharia |
2024-12-14 |
504 |
--
|
|
Gemini 3.0 vs GPT-5.1: a clear comparison for builders
|
Drishti Shah |
2025-11-19 |
921 |
--
|
|
Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller …
|
The Quill |
2023-05-07 |
237 |
--
|
|
Towards Reasoning in Large Language Models: A Survey - Summary
|
The Quill |
2023-06-09 |
248 |
--
|
|
A Survey of Large Language Models - Summary
|
The Quill |
2023-04-16 |
274 |
--
|
|
Training language models to follow instructions with human feedback - Summary
|
Rohit Agarwal |
2023-04-15 |
249 |
--
|
|
Attention Is All You Need - Summary
|
Rohit Agarwal |
2023-04-14 |
248 |
--
|
|
Expressive Text-to-Image Generation with Rich Text - Summary
|
Rohit Agarwal |
2023-04-15 |
264 |
--
|
|
The Confidence Checklist for LLMs in Production
|
Vrushank Vyas |
2023-07-01 |
51 |
--
|
|
Eight Things to Know about Large Language Models - Summary
|
The Quill |
2023-04-16 |
132 |
--
|
|
LLMs In Prod: Day 1
|
Siddharth Sambharia |
2024-12-13 |
505 |
--
|
|
Fortifying Your AI Stack: Palo Alto Networks Prisma AIRS Now on Portkey
|
Jason Roberts |
2025-08-13 |
547 |
--
|
|
AI tool sprawl: causes, risks, and how teams can regain control
|
Drishti Shah |
2025-11-20 |
1,327 |
--
|
|
June at Portkey: Agents, AI Governance, and Twenty Six Cookbooks
|
Vrushank Vyas |
2024-07-08 |
405 |
--
|
|
LLM access control in multi-provider environments
|
Drishti Shah |
2025-12-03 |
1,531 |
--
|
|
Portkey in September
|
Vrushank Vyas |
2024-10-05 |
922 |
--
|
|
GPT-4 is Getting Faster 🐇
|
Vrushank Vyas |
2023-10-16 |
257 |
--
|
|
How Snorkel evaluates and trains top AI models
|
Shae Selix |
2025-11-04 |
2,388 |
--
|
|
LoRA: Low-Rank Adaptation of Large Language Models - Summary
|
Rohit Agarwal |
2023-04-15 |
271 |
--
|
|
Open WebUI vs LibreChat: Choose the Right ChatGPT UI for Your Organization
|
Vrushank Vyas |
2025-02-19 |
1,154 |
--
|
|
Elevate Your ToolJet Experience with Portkey AI
|
Kavya MD |
2024-10-29 |
853 |
--
|
|
OpenAI DevDay's Implications for LLM Apps in Prod
|
Rohit Agarwal |
2023-11-07 |
1,041 |
--
|
|
Are We Really Making Much Progress in Text Classification? A Comparative Review …
|
The Quill |
2023-06-09 |
239 |
--
|
|
Building Production-Ready RAG Apps
|
Team Hasura |
2023-10-18 |
1,143 |
--
|
|
Launching Prompt Engineering Studio
|
Vrushank Vyas |
2025-03-17 |
840 |
--
|
|
How does an AI gateway improve building AI apps
|
Drishti Shah |
2025-12-16 |
1,527 |
--
|
|
Portkey Named a Cool Vendor in the 2025 Gartner® Cool Vendors™ in …
|
Drishti Shah |
2025-10-29 |
1,051 |
--
|
|
April Cool: 154 PR Merges
|
Rohit Agarwal |
2024-04-30 |
410 |
--
|
|
Portkey is Joining Hacktoberfest
|
Ashirwad Karande |
2024-10-09 |
339 |
--
|
|
Prompt Injection Attacks in LLMs: What Are They and How to Prevent …
|
Sabrina Shoshani |
2024-12-10 |
3,011 |
--
|
|
Generative Agents: Interactive Simulacra of Human Behavior - Summary
|
The Quill |
2023-04-16 |
283 |
--
|
|
⭐ The Developer’s Guide to OpenTelemetry: A Real-Time Journey into Observability
|
Kavya MD |
2024-10-15 |
1,124 |
--
|
|
Tracking LLM token usage across providers, teams and workloads
|
Drishti Shah |
2025-12-04 |
1,437 |
--
|
|
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, …
|
The Quill |
2024-12-26 |
471 |
--
|
|
Unpacking Semantic Caching at Walmart
|
Vrushank Vyas |
2024-02-05 |
576 |
--
|
|
Architecting for Trust: A Strategic Perspective on the MCP Registry for the …
|
Rohit Agarwal |
2025-09-09 |
793 |
--
|
|
Portkey on the AWS Marketplace
|
Siddharth Sambharia |
2025-03-16 |
572 |
--
|
|
Scaling Transformer to 1M tokens and beyond with RMT - Summary
|
The Quill |
2023-05-07 |
233 |
--
|
|
GPT Understands, Too - Summary
|
Rohit Agarwal |
2023-04-15 |
244 |
--
|
|
MiLMo:Minority Multilingual Pre-trained Language Model - Summary
|
The Quill |
2023-04-20 |
200 |
--
|
|
⭐️ Analyze your LLM calls - 2.0
|
Rohit Agarwal |
2023-08-07 |
621 |
--
|
|
Everything We Know About Claude Code Limits
|
Rohit Agarwal |
2025-07-29 |
736 |
--
|
|
Buyer’s guide to LLM observability tools 2026
|
Drishti Shah |
2025-11-27 |
1,518 |
--
|
|
SLiC-HF: Sequence Likelihood Calibration with Human Feedback - Summary
|
The Quill |
2023-05-21 |
222 |
--
|
|
Deep Dive: OpenAI's o1 - The Dawn of Deliberate AI
|
Rohit Agarwal |
2024-12-08 |
1,268 |
--
|
|
Segment Everything Everywhere All at Once - Summary
|
The Quill |
2023-04-16 |
201 |
--
|
|
AI cost observability: A practical guide to understanding and managing LLM spend
|
Drishti Shah |
2025-11-21 |
1,861 |
--
|
|
⭐ Building Reliable LLM Apps: 5 Things To Know
|
Rohit Agarwal |
2023-08-01 |
1,159 |
--
|
|
Supercharging Open-source LLMs: Your Gateway to 250+ Models
|
Vrushank Vyas |
2024-08-05 |
890 |
--
|
|
Attention Isn’t All You Need
|
Siddharth Sambharia |
2024-08-22 |
838 |
--
|
|
Generative Agents: Interactive Simulacra of Human Behavior - Summary
|
The Quill |
2023-04-16 |
197 |
--
|
|
Mixtral of Experts - Summary
|
The Quill |
2024-01-09 |
343 |
--
|
|
Portkey x Pillar - Enterprise-grade Security for LLMs in Production
|
Vrushank Vyas |
2024-08-15 |
485 |
--
|
|
Post Processing Recommender Systems with Knowledge Graphs for Recency, Popularity, and Diversity …
|
The Quill |
2023-06-02 |
221 |
--
|
|
Skeleton-of-Thought: Large Language Models Can Do Parallel Decoding - Summary
|
The Quill |
2023-08-21 |
174 |
--
|
|
Expanding the AI Gateway with Google Vertex AI Integration
|
The Quill |
2024-04-29 |
394 |
--
|
|
Anyscale's OSS Models + Portkey's Ops Stack
|
Rohit Agarwal |
2023-12-12 |
372 |
--
|
|
OpenAI - Fine-tune GPT-4o with images and text
|
Kavya MD |
2024-10-20 |
1,044 |
--
|
|
Evaluating Long-Context LLMs
|
The Quill |
2025-02-14 |
371 |
--
|
|
How to differentiate your AI product - Jasper style!
|
The Quill |
2023-09-20 |
439 |
--
|
|
August at Portkey: 2 BILLION Requests, Guardrails, Tracing, and More
|
Vrushank Vyas |
2024-09-05 |
559 |
--
|
|
Discovering Language Model Behaviors with Model-Written Evaluations - Summary
|
The Quill |
2023-05-08 |
415 |
--
|
|
LLM routing techniques for high-volume applications
|
Drishti Shah |
2025-12-05 |
1,453 |
--
|
|
Understanding MCP Authorization
|
Drishti Shah |
2025-12-17 |
1,188 |
--
|
|
What is a virtual MCP server: Need, benefits, use cases
|
Drishti Shah |
2025-12-22 |
634 |
--
|
|
Enterprise MCP access control: managing tools, servers, and agents
|
Drishti Shah |
2025-12-23 |
1,024 |
--
|
|
MCP tool discovery for autonomous LLM agents
|
Drishti Shah |
2025-12-26 |
737 |
--
|
|
OpenCode: token usage, costs, and access control
|
Drishti Shah |
2025-12-30 |
870 |
--
|
|
LLM hallucinations in production
|
Danna Wermus |
2026-01-06 |
1,147 |
--
|
|
How Fontys ICT built an institutional AI platform with a gateway architecture
|
Koen Suilen |
2026-01-10 |
1,751 |
--
|
|
We Tracked $93M in LLM Spends Last Year. Now the Data is …
|
Vrushank Vyas |
2026-01-12 |
717 |
--
|
|
The State of AI FinOps 2025: Key Insights from FinOps Foundation's Latest …
|
Vrushank Vyas |
2025-02-20 |
1,378 |
--
|
|
Benchmarking the new moderation model from OpenAI
|
Rohit Agarwal |
2024-09-27 |
1,678 |
--
|
|
Beyond Implementation: Why Audit Logs are Critical for Enterprise AI Governance
|
Vrushank Vyas |
2025-01-28 |
378 |
--
|
|
Dive into what is LLMOps
|
Vrushank Vyas |
2023-07-01 |
6,444 |
--
|
|
LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models - Summary
|
The Quill |
2023-10-14 |
230 |
--
|
|
Open Sourcing Guardrails on the Gateway Framework
|
Rohit Agarwal |
2024-08-14 |
542 |
--
|
|
Multi-LLM Text Summarization
|
The Quill |
2024-12-26 |
396 |
--
|
|
MCP primitives: the mental model behind the protocol
|
Drishti Shah |
2025-12-15 |
952 |
--
|
|
Beyond the Hype: The Enterprise AI Blueprint You Need Now (And Why …
|
Rohit Agarwal |
2025-04-08 |
1,309 |
--
|
|
Transforming E-Commerce Search with Semantic Cache: Insights from Walmart's Journey
|
Rohit Agarwal |
2024-02-09 |
861 |
--
|
|
OpenAI's New Agent Tools: Navigating Strategic Implications for Enterprise AI
|
Vrushank Vyas |
2025-03-12 |
1,742 |
--
|
|
Jamming on Event-Driven Architecture and MCP for Multi-Agentic Systems
|
Siddharth Sambharia |
2025-01-13 |
530 |
--
|
|
Introducing the MCP Gateway
|
Rohit Agarwal |
2026-01-21 |
1,055 |
--
|
|
Securing the MCP Gateway: Lasso Partners with Portkey to Deliver Enterprise-Grade Agentic …
|
Eliran Suisa |
2026-02-16 |
908 |
--
|
|
Portkey Raises $15M Series A to Scale the Unified Control Plane for …
|
Rohit Agarwal |
2026-02-19 |
698 |
--
|
|
Open AI Responses API vs. Chat Completions vs. Anthropic Messages API
|
Drishti Shah |
2026-02-23 |
1,377 |
--
|
|
The best approach to compare LLM outputs
|
Drishti Shah |
2026-02-24 |
1,143 |
--
|
|
How to host an AI Hackathon without losing control of your keys …
|
Drishti Shah |
2026-02-25 |
1,326 |
--
|
|
LLM Deployment Pipeline Explained Step by Step
|
Rebecca McCandler |
2026-02-27 |
1,845 |
--
|
|
Claude Code agents: what they are, how they work, and how to …
|
Drishti Shah |
2026-03-09 |
1,096 |
--
|
|
Claude Code best practices for enterprise teams
|
Drishti Shah |
2026-03-10 |
1,336 |
--
|
|
MCP vs RAG Compared for Production Teams
|
Drishti Shah |
2026-03-12 |
1,686 |
--
|
|
What Makes Enterprise LLMs Different from General-Purpose AI Tools
|
Rebecca McCandler |
2026-03-13 |
1,699 |
--
|
|
1 Trillion Tokens and the Death of the Chatbot
|
Rohit Agarwal |
2026-03-19 |
1,194 |
--
|
|
MCP vs Function Calling – How They Actually Work Together
|
Drishti Shah |
2026-03-20 |
1,979 |
--
|
|
GPT-5.4 vs Claude Opus 4.6: a guide to choosing the right model
|
Drishti Shah |
2026-03-23 |
1,777 |
--
|
|
The Gateway Grew Up
|
Swetha Sridhar |
2026-03-24 |
600 |
--
|
|
Enterprise AI Architecture From Pilot to Production
|
Vrushank Vyas |
2026-03-23 |
2,900 |
--
|
|
What is AI lifecycle management?
|
Drishti Shah |
2026-03-24 |
1,138 |
--
|
|
Akto Partners with Portkey to Bring Guardrails to AI Gateway
|
Drishti Shah |
2026-03-31 |
891 |
--
|
|
What is autoinstrumentation?
|
Drishti Shah |
2026-04-01 |
1,061 |
--
|
|
Stop hardcoding API keys in your AI apps
|
Drishti Shah |
2026-04-02 |
980 |
--
|
|
Rate limiting for LLM applications: Why it matters and how to implement …
|
Drishti Shah |
2026-04-06 |
1,375 |
--
|
|
Tool Provisioning in MCP Servers: Controlling AI Agent Access in Production
|
Siddharth Sambharia |
2026-04-07 |
1,633 |
--
|
|
Cursor best practices for enterprise teams
|
Drishti Shah |
2026-04-09 |
1,515 |
--
|
|
The Harness Tax: The Dead Weight Inside Your Coding Agent
|
Siddharth Sambharia |
2026-04-13 |
861 |
--
|
|
LLM pricing is 100x harder than you think
|
Siddharth Sambharia |
2026-04-15 |
1,395 |
--
|
|
Moving Fast Has a Security Bill and It Just Came Due
|
Rohit Agarwal |
2026-04-16 |
1,749 |
--
|
|
What is AIOps?
|
Drishti Shah |
2026-04-16 |
1,175 |
--
|
|
Conductor × Portkey is now live
|
Swetha Sridhar |
2026-04-17 |
452 |
--
|
|
How to choose the right AIOps platform
|
Drishti Shah |
2026-04-17 |
868 |
--
|
|
Semantic caching thresholds and why they matter
|
Swetha Sridhar |
2026-04-18 |
2,267 |
--
|
|
AI Agent governance
|
Drishti Shah |
2026-04-19 |
1,058 |
--
|
|
OpenAI Codex best practices
|
Drishti Shah |
2026-04-20 |
1,058 |
--
|
|
n8n Best Practices
|
Drishti Shah |
2026-04-21 |
1,427 |
--
|
|
Introducing the Agent Gateway
|
Rohit Agarwal |
2026-04-21 |
440 |
--
|
|
Your First AI Agent Will Go Fine. Your Fiftieth Is Where Things …
|
Swetha Sridhar |
2026-04-22 |
1,405 |
--
|
|
Who owns Claude Code at your company? A platform team's guide to …
|
Siddharth Sambharia |
2026-04-24 |
1,234 |
--
|
|
Introducing Skills Registry
|
Siddharth Sambharia |
2026-04-23 |
503 |
--
|
|
GitHub Copilot best practices for teams
|
Drishti Shah |
2026-04-27 |
1,035 |
--
|
|
What is AgentOps?
|
Drishti Shah |
2026-04-29 |
1,208 |
--
|
|
What’s an agent gateway?
|
Drishti Shah |
2026-05-03 |
1,266 |
--
|
|
What MCP Governance Actually Means in Production
|
Swetha Sridhar |
2026-05-24 |
1,821 |
--
|
|
Why Every Agent Vulnerability is a Trust Boundary Failure
|
Narendranath Gogineni |
2026-06-28 |
1,936 |
--
|
|
Why financial firms need granular governance for Gen AI
|
Drishti Shah |
2025-02-21 |
931 |
--
|
|
What is an AI Gateway and how to choose one in 2026
|
Drishti Shah |
2025-09-05 |
1,530 |
--
|
|
ReAct: Synergizing Reasoning and Acting in Language Models - Summary
|
The Quill |
2023-04-27 |
267 |
--
|
|
How to design a reliable fallback system for LLM apps using an …
|
Drishti Shah |
2025-06-06 |
1,394 |
--
|
|
Bridging the Gap: How Portkey AI Gateway Connected with MongoDB Helps Productionize …
|
Siddharth Sambharia |
2024-09-01 |
1,908 |
--
|
|
Real-world applications and examples of AI agents
|
Drishti Shah |
2025-02-06 |
1,174 |
--
|
|
How to scale AI apps - Lessons from building a billion-scale AI …
|
Drishti Shah |
2024-12-06 |
676 |
--
|
|
Large Language Models Are Human-Level Prompt Engineers - Summary
|
Rohit Agarwal |
2023-04-15 |
256 |
--
|
|
What is LLM Orchestration?
|
Drishti Shah |
2025-02-25 |
1,161 |
--
|
|
Why Portkey is the right AI Gateway for you
|
Drishti Shah |
2024-12-04 |
974 |
--
|
|
Portkey & Patronus - Bringing Responsible LLMs in Production
|
Vrushank Vyas |
2024-08-14 |
403 |
--
|
|
Building AI agent workflows with the help of an MCP gateway
|
Drishti Shah |
2025-05-27 |
905 |
--
|
|
What are AI agents?
|
Drishti Shah |
2025-01-14 |
1,389 |
--
|
|
From standard to ecosystem: the new MCP updates, Nov 2025
|
Drishti Shah |
2025-11-16 |
734 |
--
|
|
LLM cost attribution: Tracking and optimizing spend for GenAI apps
|
Drishti Shah |
2025-04-29 |
928 |
--
|
|
LibreChat vs Open WebUI: Choose the Right ChatGPT UI for Your Organization
|
Siddharth Sambharia |
2024-11-13 |
1,162 |
--
|
|
The hidden challenge of MCP adoption in enterprises in 2025
|
Drishti Shah |
2025-08-25 |
729 |
--
|
|
Claude Skills: definition, use cases, and limitations
|
Drishti Shah |
2025-11-05 |
1,521 |
--
|
|
AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head - Summary
|
The Quill |
2023-05-06 |
325 |
--
|
|
How to improve LLM performance
|
Drishti Shah |
2025-01-27 |
734 |
--
|
|
The real cost of building an LLM gateway
|
Drishti Shah |
2025-01-03 |
1,265 |
--
|
|
🌖 Announcing $3M Seed Round to Bring LLMs to Production
|
Rohit Agarwal |
2023-08-23 |
1,096 |
--
|
|
Why We Chose TypeScript Over Python for the World's Fastest AI Gateway
|
Ayush |
2024-09-03 |
946 |
--
|
|
OpenAI’s Prompt Caching: A Deep Dive
|
Ayush |
2024-10-20 |
1,018 |
--
|
|
How a model catalog accelerates LLM development
|
Drishti Shah |
2025-06-27 |
906 |
--
|
|
What is LLM tool calling, and how does it work?
|
Drishti Shah |
2025-04-12 |
830 |
--
|
|
Powering GenAI initiatives in insurance companies to go from pilot to production
|
Drishti Shah |
2025-10-28 |
625 |
--
|
|
Bring Your Agents to Production with Portkey
|
Siddharth Sambharia |
2024-08-21 |
991 |
--
|
|
The Power of Scale for Parameter-Efficient Prompt Tuning - Summary
|
Rohit Agarwal |
2023-04-15 |
357 |
--
|
|
How to balance AI model accuracy, performance, and costs with an AI …
|
Vrushank Vyas |
2025-06-25 |
1,469 |
--
|
|
MCP Message Types: Complete MCP JSON-RPC Reference Guide
|
Rohit Agarwal |
2025-09-22 |
2,555 |
--
|
|
What is an LLM Gateway?
|
Ayush |
2024-11-26 |
1,648 |
--
|
|
SegGPT: Segmenting Everything In Context - Summary
|
The Quill |
2023-04-16 |
247 |
--
|
|
What is Automatic Prompt Engineering?
|
Drishti Shah |
2024-10-09 |
2,013 |
--
|
|
Scaling LibreChat for enterprise use: tracking, visibility, and governance
|
Siddharth Sambharia |
2025-07-21 |
945 |
--
|
|
Breaking down the real cost factors behind generative AI
|
Drishti Shah |
2025-04-17 |
1,079 |
--
|
|
Claude Sonnet 4.5 vs GPT-5: performance, efficiency, and pricing compared.
|
Drishti Shah |
2025-10-01 |
1,430 |
--
|
|
MRKL Systems: A modular, neuro-symbolic architecture that combines large language models, external …
|
The Quill |
2023-04-27 |
319 |
--
|
|
AI Prompts for Sales Reps
|
Drishti Shah |
2025-03-17 |
1,367 |
--
|
|
Model Context Protocol for building reliable, enterprise LLM applications
|
Drishti Shah |
2024-12-16 |
973 |
--
|
|
What is AI governance?
|
Drishti Shah |
2025-02-03 |
1,164 |
--
|
|
AI Gateway for governance in Azure AI apps
|
Drishti Shah |
2025-06-10 |
856 |
--
|
|
Reducing AI hallucinations with guardrails
|
Drishti Shah |
2025-01-09 |
997 |
--
|
|
⭐️ Getting Started with Llama 2
|
Rohit Agarwal |
2023-09-26 |
1,773 |
--
|
|
Real-time Guardrails vs Batch Evals: Understanding Safety Mechanisms in LLM Applications
|
Ayush |
2024-12-12 |
404 |
--
|
|
End-to-End Debugging: Tracing Failures from the LLM Call to the User Experience
|
Drishti Shah |
2025-09-15 |
1,285 |
--
|
|
Why LLM security is non-negotiable
|
Drishti Shah |
2025-06-03 |
720 |
--
|
|
Best practices for securing and governing tool access in an MCP hub
|
Drishti Shah |
2025-08-15 |
927 |
--
|
|
How AI gateways enable scalable agent orchestration
|
Drishti Shah |
2025-08-21 |
973 |
--
|
|
Build vs Buy - LLM Gateways
|
Drishti Shah |
2024-12-10 |
1,175 |
--
|
|
Why reliability in AI applications is now a competitive differentiator
|
Drishti Shah |
2025-09-10 |
1,214 |
--
|
|
AI Prompts for Product Marketers
|
Drishti Shah |
2025-03-20 |
1,193 |
--
|
|
Why Multi-LLM Provider Support is Critical for Enterprises
|
Drishti Shah |
2025-02-12 |
1,115 |
--
|
|
Using an MCP (Model Context Protocol) gateway to unify context across multi-step …
|
Drishti Shah |
2025-05-23 |
775 |
--
|
|
Retries, fallbacks, and circuit breakers in LLM apps: what to use when
|
Drishti Shah |
2025-07-17 |
985 |
--
|
|
Building a Robust RAG App with Guardrails using Portkey and MongoDB
|
Siddharth Sambharia |
2024-09-01 |
923 |
--
|
|
Making Claude Code work for enterprise-scale use with an AI Gateway
|
Drishti Shah |
2025-07-19 |
725 |
--
|
|
What is Knowledge Augmented Generation (KAG)?
|
Drishti Shah |
2025-01-21 |
1,353 |
--
|
|
7 Ways to Make Your Vercel AI SDK App Production-Ready
|
Siddharth Sambharia |
2024-10-08 |
1,405 |
--
|
|
Prompt engineering vs. fine-tuning: What’s better for your use case?
|
Drishti Shah |
2025-02-17 |
1,053 |
--
|
|
Prompting Claude 3.5 vs 3.7
|
Drishti Shah |
2025-03-29 |
1,081 |
--
|
|
Prompt engineering for low-resource languages
|
Drishti Shah |
2025-02-11 |
1,004 |
--
|
|
Scaling and managing LLM applications: The essential guide to LLMOps tools
|
Drishti Shah |
2025-04-23 |
1,483 |
--
|
|
Make Cline enterprise-ready using an AI Gateway
|
Drishti Shah |
2025-06-26 |
912 |
--
|
|
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models - Summary
|
Rohit Agarwal |
2023-04-15 |
332 |
--
|
|
Bringing multimodal models to production with an AI Gateway
|
Drishti Shah |
2025-06-12 |
1,019 |
--
|
|
Prompt Engineering for Stable Diffusion
|
Drishti Shah |
2025-03-12 |
1,737 |
--
|
|
AI Gateway vs API Gateway - What's the difference
|
Drishti Shah |
2024-11-29 |
1,350 |
--
|
|
Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models - Summary
|
The Quill |
2023-04-20 |
528 |
--
|
|
LLM Observability is now a business function for AI
|
Drishti Shah |
2025-10-25 |
1,065 |
--
|
|
How to add enterprise controls to OpenWebUI: cost tracking, access control, and …
|
Siddharth Sambharia |
2025-07-17 |
779 |
--
|
|
How to Optimize Token Efficiency When Prompting
|
Drishti Shah |
2025-03-21 |
936 |
--
|
|
FinOps practices to optimize GenAI costs and maximize efficiency
|
Ayush |
2024-11-07 |
946 |
--
|
|
What is AI interoperability, and why does it matter in the age …
|
Drishti Shah |
2025-05-12 |
1,075 |
--
|
|
How to implement budget limits and alerts in LLM applications
|
Drishti Shah |
2025-05-15 |
1,154 |
--
|
|
Managing and deploying prompts at scale without breaking your pipeline
|
Drishti Shah |
2025-07-07 |
873 |
--
|
|
Self-Consistency Improves Chain of Thought Reasoning in Language Models - Summary
|
Rohit Agarwal |
2023-04-15 |
301 |
--
|
|
Expanding AI safety with Qualifire guardrails on Portkey
|
Danna Wermus |
2025-11-17 |
469 |
--
|
|
MCP hub vs MCP registry: What’s the difference?
|
Drishti Shah |
2025-08-13 |
712 |
--
|
|
Task-Based LLM Routing: Optimizing LLM Performance for the Right Job
|
Drishti Shah |
2025-04-08 |
1,329 |
--
|
|
How an AI gateway improves the management of AI deployments
|
Drishti Shah |
2025-06-11 |
1,006 |
--
|
|
LLMs in Prod 2025: Insights from 2 Trillion+ Tokens
|
Siddharth Sambharia |
2025-01-21 |
419 |
--
|
|
Model Context Protocol (MCP): Everything You Need to Know to Get Started
|
Siddharth Sambharia |
2025-03-10 |
1,259 |
--
|
|
Role-based access control (RBAC) for LLM applications
|
Drishti Shah |
2025-05-29 |
978 |
--
|
|
Prompt Security and Guardrails: How to Ensure Safe Outputs
|
Drishti Shah |
2024-11-14 |
1,358 |
--
|
|
How to secure your entire LLM lifecycle
|
Drishti Shah |
2025-06-05 |
692 |
--
|
|
Evaluating Prompt Effectiveness: Key Metrics and Tools
|
Drishti Shah |
2024-11-01 |
1,833 |
--
|
|
Partnering with F5 to Productionize Enterprise AI
|
Rohit Agarwal |
2024-05-02 |
685 |
--
|
|
Accelerating LLMs with Skeleton-of-Thought Prompting
|
Drishti Shah |
2025-03-25 |
1,436 |
--
|
|
GPT 5 vs Claude 4
|
Drishti Shah |
2025-09-04 |
1,505 |
--
|
|
Using metadata for better LLM observability and debugging
|
Drishti Shah |
2025-05-14 |
870 |
--
|
|
How LLM tracing helps you debug and optimize GenAI apps
|
Drishti Shah |
2025-05-01 |
927 |
--
|
|
MCP vs A2A
|
Drishti Shah |
2025-05-05 |
914 |
--
|
|
AI Prompts for Social Media Marketers
|
Drishti Shah |
2025-03-14 |
1,327 |
--
|
|
Securing your AI via AI Gateways
|
Drishti Shah |
2025-04-14 |
605 |
--
|
|
Debugging agent workflows with MCP observability
|
Drishti Shah |
2025-06-09 |
758 |
--
|
|
What are MCP connectors?
|
Drishti Shah |
2025-08-04 |
744 |
--
|
|
Lifecycle of a Prompt
|
Drishti Shah |
2025-02-27 |
1,067 |
--
|
|
The complete guide to LLM observability for 2026
|
Drishti Shah |
2025-11-04 |
2,440 |
--
|
|
Making OpenAI's Typescript SDK production-ready with an AI Gateway
|
Drishti Shah |
2025-06-17 |
1,094 |
--
|
|
Portkey Goes Multimodal
|
Rohit Agarwal |
2024-03-28 |
1,062 |
--
|
|
Comparing lean LLMs: GPT-5 Nano and Claude Haiku 4.5
|
Drishti Shah |
2025-10-16 |
1,131 |
--
|
|
Just Tell Me: Prompt Engineering in Business Process Management - Summary
|
The Quill |
2023-04-22 |
210 |
--
|
|
Build resilient Azure AI applications with an AI Gateway
|
Drishti Shah |
2025-05-15 |
1,106 |
--
|
|
⭐️ OpenAI Model Deprecation Guide
|
Vrushank Vyas |
2024-01-02 |
338 |
--
|
|
LLM observability vs monitoring
|
Drishti Shah |
2025-01-01 |
1,563 |
--
|
|
⭐️ Implementing FrugalGPT: Reducing LLM Costs & Improving Performance
|
Rohit Agarwal |
2024-04-22 |
2,813 |
--
|
|
What is shadow AI, and why is it a real risk for …
|
Drishti Shah |
2025-07-11 |
1,141 |
--
|
|
Challenges Agentic AI Companies Face in Enterprise Adoption
|
Drishti Shah |
2025-02-26 |
1,102 |
--
|
|
Why forward compatibility is critical for Agentic AI companies
|
Drishti Shah |
2025-04-16 |
729 |
--
|
|
⭐️ Decoding OpenAI Evals
|
Rohit Agarwal |
2023-05-10 |
2,316 |
--
|
|
Securing enterprise AI with gateways and guardrails
|
Tom Prenderville |
2025-10-01 |
1,641 |
--
|
|
Geo-location based LLM routing: Why it matters and how to do it …
|
Drishti Shah |
2025-04-09 |
898 |
--
|
|
How MCP(Model Context Protocol) handles context management in high-throughput scenarios
|
Drishti Shah |
2025-03-24 |
1,379 |
--
|
|
How to Build Multi-Agent AI Systems with OpenAI Swarm & Secure Them …
|
Drishti Shah |
2024-11-01 |
266 |
--
|
|
Mastering role prompting: How to get the best responses from LLMs
|
Drishti Shah |
2025-03-03 |
970 |
--
|
|
Why enterprises need to rethink how employees access LLMs
|
Drishti Shah |
2025-07-09 |
689 |
--
|
|
Meta prompting: Enhancing LLM Performance
|
Drishti Shah |
2025-03-13 |
822 |
--
|
|
Delimiters in Prompt Engineering
|
Drishti Shah |
2025-03-01 |
909 |
--
|
|
What are AI guardrails?
|
Drishti Shah |
2025-01-06 |
1,823 |
--
|
|
The AI governance problem in higher education and how to solve it
|
Drishti Shah |
2025-07-31 |
1,135 |
--
|
|
Building the world's fastest AI Gateway - stream transformers
|
Narendranath Gogineni |
2025-07-16 |
618 |
--
|
|
Using OpenAI AgentKit with Anthropic, Gemini and other providers
|
Drishti Shah |
2025-10-08 |
1,070 |
--
|
|
AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated Prompts - Summary
|
Rohit Agarwal |
2023-04-15 |
250 |
--
|
|
Failover routing strategies for LLMs in production
|
Drishti Shah |
2025-09-18 |
1,427 |
--
|
|
Understanding prompt engineering parameters
|
Drishti Shah |
2025-03-10 |
900 |
--
|
|
Top 10 MCP Servers
|
Drishti Shah |
2024-12-19 |
502 |
--
|
|
DSPy in Production
|
Ganaraj Permunda |
2024-09-03 |
950 |
--
|
|
What a modern LLMOps stack looks like in 2025
|
Drishti Shah |
2025-04-21 |
1,143 |
--
|
|
Chain-of-Thought (CoT) Capabilities in O1-mini and O1-preview
|
Ayush |
2024-10-29 |
1,829 |
--
|
|
Portkey Prompt Engineering Studio User-Centered Design Case Study
|
Nikhil Kanda |
2025-03-25 |
2,785 |
--
|
|
Simplifying MCP server authentication for enterprises
|
Drishti Shah |
2025-08-27 |
966 |
--
|
|
Syngenta: Driving AI adoption through hackathons
|
Rahul Miragi |
2025-10-09 |
680 |
--
|
|
From Arm Pain to AI Gateway: Why I Chose Portkey for Managing …
|
Rahul Bansal |
2025-10-06 |
969 |
--
|
|
The Evolution from AI Assistants to AI Agents
|
Drishti Shah |
2025-02-13 |
1,049 |
--
|
|
How to use Claude Code with Bedrock, Vertex AI and Anthropic
|
Drishti Shah |
2025-08-10 |
684 |
--
|
|
Using Prompt Chaining for Complex Tasks
|
Ayush |
2024-11-08 |
1,191 |
--
|
|
Zero-Shot vs. Few-Shot Prompting: Choosing the Right Approach for Your AI Model
|
Drishti Shah |
2024-10-27 |
1,400 |
--
|
|
Load balancing in multi-LLM setups: Techniques for optimal performance
|
Drishti Shah |
2025-02-19 |
1,061 |
--
|
|
Building Production-Ready AI Security: How Falco Vanguard Solves Enterprise Security Challenges with …
|
Miguel De Los Santos |
2025-09-09 |
1,972 |
--
|
|
What is LLM Observability?
|
Drishti Shah |
2024-11-11 |
890 |
--
|
|
Boosted Prompt Ensembles for Large Language Models -Summary
|
The Quill |
2023-05-17 |
375 |
--
|
|
ChatGPT vs DeepSeek vs Claude - which LLM Fits Your Needs?
|
Drishti Shah |
2025-02-01 |
2,712 |
--
|
|
Claude Code vs Cursor: What to choose?
|
Drishti Shah |
2025-08-06 |
1,904 |
--
|
|
How to scale GenAI apps built on Azure AI services
|
Drishti Shah |
2025-05-07 |
1,174 |
--
|
|
What is an MCP hub and why it’s the next layer for …
|
Drishti Shah |
2025-08-11 |
629 |
--
|
|
Bringing GenAI to the classroom
|
Drishti Shah |
2025-04-13 |
1,087 |
--
|
|
AI Governance Checklist for 2025: control and safety via an AI gateway
|
Drishti Shah |
2025-06-24 |
2,712 |
--
|
|
Basic AI Prompts for Developers: Practical Examples for Everyday Tasks
|
Drishti Shah |
2025-03-27 |
851 |
--
|
|
⭐ Semantic Cache for Large Language Models
|
Vrushank Vyas |
2023-07-11 |
1,711 |
--
|
|
Ethical considerations and bias mitigation in AI
|
Drishti Shah |
2025-04-04 |
736 |
--
|
|
How Stardog uses NVIDIA's Triton Server to build scalable, enterprise-grade AI applications
|
Tapan Sharma |
2024-11-29 |
463 |
--
|
|
How to identify and mitigate shadow AI risks in organizations using an …
|
Drishti Shah |
2025-07-14 |
1,278 |
--
|
|
⭐️ Ranking LLMs with Elo Ratings
|
Rohit Agarwal |
2023-04-18 |
1,231 |
--
|
|
Why connecting OTel traces with LLM logs is critical for agent workflows
|
Siddharth Sambharia |
2025-08-23 |
775 |
--
|
|
Canary Testing for LLM Apps
|
Drishti Shah |
2025-04-05 |
752 |
--
|
|
FinOps chargeback and how it can help GenAI platforms
|
Drishti Shah |
2025-04-02 |
725 |
--
|
|
COSTAR Prompt Engineering: What It Is and Why It Matters
|
Drishti Shah |
2025-03-07 |
1,146 |
--
|
|
LLM Grounding: How to Keep AI Outputs Accurate and Reliable
|
Drishti Shah |
2025-02-28 |
953 |
--
|
|
Benefits of using MCP over traditional integration methods
|
Drishti Shah |
2025-03-29 |
540 |
--
|
|
Tackling rate limiting for LLM apps
|
Drishti Shah |
2025-01-24 |
1,219 |
--
|
|
Opentelemetry semantic conventions for GenAI traces
|
Narendranath Gogineni |
2025-10-29 |
343 |
--
|
|
Prompting Chatgpt vs Claude
|
Drishti Shah |
2024-11-21 |
1,704 |
--
|
|
Sparks of Artificial General Intelligence: Early experiments with GPT-4 - Summary
|
The Quill |
2023-04-20 |
398 |
--
|
|
The hidden technical debt in LLM apps
|
Drishti Shah |
2025-04-24 |
852 |
--
|
|
Prompt engineering techniques for effective AI outputs
|
Drishti Shah |
2024-12-24 |
2,860 |
--
|
|
LLM proxy vs AI gateway: what’s the difference and which one do …
|
Drishti Shah |
2025-07-10 |
1,321 |
--
|
|
The Complete Guide to Prompt Engineering
|
Drishti Shah |
2024-10-26 |
2,121 |
--
|
|
Simplifying LLM batch inference
|
Mahesh Vagicherla |
2025-08-22 |
955 |
--
|
|
Orchestrating multiple MCP servers in a single AI workflow
|
Drishti Shah |
2025-08-18 |
1,256 |
--
|
|
How to use Roo Code in your organization (the right way)
|
Drishti Shah |
2025-06-16 |
1,097 |
--
|
|
Scaling production AI: Cerebras joins the Portkey ecosystem
|
Drishti Shah |
2025-09-09 |
480 |
--
|
|
What is AI TRiSM?
|
Drishti Shah |
2025-04-18 |
1,429 |
--
|
|
The most reliable AI gateway for production systems
|
Drishti Shah |
2025-10-06 |
1,456 |
--
|