Submit event
Webinar on Inside LLM Context Windows, Token Consumption, and Practical Context Engineering by Toni Ramchandani

Webinar on Inside LLM Context Windows, Token Consumption, and Practical Context Engineering by Toni Ramchandani

26 Aug 202617:00 - 18:30 Asia/KolkataOnline0 AttendeesOpen

Description

We're excited to announce the upcoming TTTribeCast Webinar on: "Inside LLM Context Windows, Token Consumption, and Practical Context Engineering" by Toni Ramchandani ( Vice President at MSCI Inc.) 📅 26th August 2026 | Wednesday 🕘 5:00 PM – 6:30 PM IST 💻 Live on Zoom | Free to attend What will Toni speak about: This session explores what actually happens inside an LLM when context enters the model, covering tokenization, embeddings, attention, QKV, prefill, decode, and KV caching. It examines how context windows are allocated and consumed, and why simply increasing context length does not always improve model performance, with practical examples from Lost in the Middle, RULER, and NoLiMa. The session then moves into practical context engineering, covering signal density, token reduction, and how to identify the minimum sufficient context for a task, followed by an optimization ladder for improving local LLM performance with Ollama. You'll walk away with: ✅ Understand how tokens, attention, and KV caching shape LLM context and performance. ✅ See why larger context windows don’t always mean better model performance.✅ Learn practical techniques to reduce token waste and preserve high-value context. ✅ Optimize local LLM performance by balancing context, memory, latency, and quality. This session is for you if you’re: • A QA, engineering, or technology leader looking to understand how context impacts LLM performance and cost • A QA manager, engineering manager, or architect evaluating practical AI adoption and optimization strategies • An engineering professional working with LLMs and looking to improve context efficiency, reliability, and performance • Someone making decisions around AI infrastructure, local LLMs, or context engineering for real-world use cases About the speaker: Toni Ramchandani is the Vice President of Applied AI Engineering at MSCI Inc., bringing over 17 years of experience in AI, data, cloud, and engineering execution. He specializes in enterprise AI agents, RAG, LLM evaluation, and AI observability to turn emerging tech into scalable, production-ready enterprise capabilities. Toni is also an author, having written A Generative Journey to AI, and an avid endurance athlete who regularly runs, rides, and hikes. A few things people usually ask: Become a Speaker at The Test Tribe Events Wish to share your insights with the community at a future event? Submit your talk idea here, and if it resonates with the perspective we are trying to build at our events, our team will connect with you. About The Test Tribe: The Test Tribe, established in 2018, is the world's largest software testing community and an EdTech startup. We empower testing professionals globally with expert courses, memberships, events, and community initiatives, fostering collaboration, learning, and growth. With over 925+ events and a global reach of 170K+ testers from 130+ countries, our mission is to provide life-altering growth to every testing professional.

Let your network know you`re going

Share this event to start conversations, invite colleagues, and connect before it begins.

Related events

Artificial Intelligence Platform

Free