Optimize Azure OpenAI Applications with Semantic Caching

Introduction
One of the ways to optimize cost and performance of Large Language Models (LLMs) is to cache the responses from LLMs, this is sometimes referred to as “semantic caching”. In this blog, we will discuss the approaches, benefits, common scena… Continue reading Optimize Azure OpenAI Applications with Semantic Caching

Upcoming April 2024 Microsoft 365 Champion Community Call

Join our next community call on April 23 to explore Microsoft Teams Communities. We will be starting the call at 5 minutes past the hour for both of our sessions (at 8:05 AM and 5:05 PM PT), and it will still end at the top of the hour (9:00 AM and 6:… Continue reading Upcoming April 2024 Microsoft 365 Champion Community Call

How to Customize an LLM: A Deep Dive to Tailoring an LLM for Your Business

Introduction 

 
In the world of large language models, model customization is key. It’s what transforms a standard model into a powerful tool tailored to your business needs. 
Let’s explore three techniques to customize a Large Language… Continue reading How to Customize an LLM: A Deep Dive to Tailoring an LLM for Your Business