Our Sessions

Have a look at our sessions for the next edition of Haystack Europe on September 15-16th in Berlin, Germany. If you like what you see, you can get a ticket here.

Filter by:

Track

No tracks available

Session Type
Organisation

What Gets Retrieved Gets Reinforced: Gender Bias in Agents

Talk
16. September 2026, 02:35 pm - 02:55 pm
Main Stage
If your AI agent treats Sarah and Michael differently for the same query, would you know? We built a two-layer detector, learned why retrieval counting drowns in catalog noise, and landed on counterfactual testing as the only reliable signal. You will leave with a concrete, practical approach to testing gender bias in your own agentic systems.

Training 2: LLMs as Judges for Search Result Quality

Training
14. September 2026, 01:45 pm - 05:45 pm
Main Stage
Large Language Models (LLMs) transform how we build and evaluate search systems, it’s crucial to understand how to use them effectively as “judges.” This condensed, hands-on training introduces the principles and practical techniques for implementing “LLM as a Judge” to evaluate search result quality.

Agent Memory that works with your existing Search Stack

Talk
16. September 2026, 04:00 pm - 04:45 pm
Main Stage
An agent that forgets you between sessions breaks a core UX promise. For our Agent Studio, we built memory so that it works across classic search, vectors, or both, based on what you already have. I'll share our journey: what worked, what surprised us, and why 65% is sometimes the best score you can get.

Training 1: Building Modern Search Platforms for Humans & AI

Training
14. September 2026, 09:00 am - 01:00 pm
Main Stage
Search is no longer just a product feature — it’s becoming a foundational platform that powers AI assistants, copilots, personalization engines, and internal knowledge systems. Shifting from a solid onsite search experience to a true search platform raises new questions: how should teams collaborate, what data needs to be exposed and in what form, and how do you design for both human users and AI agents at once?

Closing Note

Organizational Session
15. September 2026, 05:35 pm - 05:45 pm
Main Stage
Join us as we wrap up the first conference day!

Welcome Back

Organizational Session
16. September 2026, 09:00 am - 09:15 am
Main Stage
Join us as we kick off the second conference day!

Keynote

Talk
15. September 2026, 09:15 am - 10:00 am
Main Stage
Speaker & Topic to be announced.

Closing Notes

Organizational Session
16. September 2026, 04:45 pm - 05:00 pm
Main Stage
Join us as we wrap up Haystack Europe!

Lightning Talks

Talk
15. September 2026, 04:50 pm - 05:35 pm
Main Stage
Join us for our lightning talk session!

Introduction & Welcome

Organizational Session
15. September 2026, 09:00 am - 09:15 am
Main Stage
Join us as we kick off Haystack Europe!

Sparse encoders: bridging products and knowledge graphs

Talk
16. September 2026, 10:00 am - 10:45 am
Main Stage
Discover how Leroy Merlin bridges product catalogs with Knowledge Graphs using Sparse Encoders. We’ll share how we injected KG concepts as model tokens, tuned custom layers for 99%+ sparsity, and achieved explainable, low-latency (60ms) semantic search indexing ready for Elasticsearch.

Who Gets to Decide? Designing Agency Back Into AI Search

Talk
16. September 2026, 02:15 pm - 02:35 pm
Main Stage
Drawing on behavioral UX, cognitive offloading and spatial thinking, Öykü will share practical design principles for making AI-powered search easier to understand, inspect and challenge: How can interfaces surface sources, uncertainty and alternative perspectives while helping people remain active participants in the decision-making process.

Search (Re)platforming for Humans and AI

Talk
15. September 2026, 03:15 pm - 04:00 pm
Main Stage
Balancing business rules, vector search, and AI evaluation is a shared challenge. Modern search systems must now serve both human and agentic users. In this talk we share grounded insights on search replatforming, API-driven tools, deterministic query governance, and the growing complexities of evaluation in the age of generative AI.

Beyond RAG: Recursive Agentic Retrieval for Enterprise AI

Talk
15. September 2026, 02:15 pm - 03:00 pm
Main Stage
Think larger context window solves enterprise AI? Think again. Learn how recursive agentic retrieval progressively constructs the right context for every reasoning step, enabling AI agents to solve complex enterprise tasks beyond traditional RAG.

From Tries to Transformers: Scaling Semantic Search

Talk
15. September 2026, 11:45 am - 12:30 pm
Main Stage
Netflix search has evolved from simple Trie-based lookups to complex, intent-driven conversational queries. This talk explores our journey to integrate deep Query Understanding and semantic retrieval at scale.

Our Journey to Multi-Tenant Learned Re-Ranking for Commerce

Talk
15. September 2026, 04:00 pm - 04:45 pm
Main Stage
Search is essential to how shoppers discover products, and we asked whether machine-learned re-ranking could improve search effectiveness across hundreds of commerce clients. This is that journey: from a baseline and a null A/B test, through modelling, architecture, and evaluation, to a client revenue win, and toward mass generalisation.

Frankensteining LLMs

Talk
15. September 2026, 01:30 pm - 02:15 pm
Main Stage
TNG's R1T Chimera models reached over 10 billion daily tokens on OpenRouter. How can a small consultancy produce its own LLMs? In our case: by Frankensteining them. This talk motivates: you can adapt LLMs to your own needs without being a big research lab. Including theory, technicalities, and lessons learned.

Teaching Your Shopping Assistant to Learn From Its Mistakes

Talk
16. September 2026, 11:45 am - 12:30 pm
Main Stage
At OTTO, we want conversational AI to become the primary interface for product discovery. We built a weekly loop that mines production sessions, grounds an LLM judge in product discovery signals, clusters failures, and ships fixes to search tools and orchestration. In this session, you'll get the building blocks, our key findings and outcomes.

Query Understanding with Late Interaction & Wormhole Vectors

Talk
16. September 2026, 09:15 am - 10:00 am
Main Stage
Modern IR techniques like Late Interaction (multivector representations) and wormhole vectors (hopping between sparse & dense vector spaces) newly expand our toolbox for building state of the art Query Understanding. We'll show you how, including ranking benchmarks that blow away today's popular hybrid search (BM25 + vector similarity) techniques.

The RAG Cost Curve: When an index beats live search

Talk
16. September 2026, 01:30 pm - 02:15 pm
Main Stage
"Just let the agent search live" has a cost curve, so does a compressed index. The curves cross but most teams pick a side before doing the maths. We'll use open benchmarks on quantisation and refinement to prove what compression really costs you in recall, in answer quality, and in dollars against live search.

Hyperbolic embedding models for visual search

Talk
15. September 2026, 11:00 am - 11:45 am
Main Stage
We trained an open image-text embedding model in hyperbolic space, so visual search can use hierarchy as well as similarity. The talk shows how radius and aperture create hierarchy inside the embedding model, how that helps retrieve specific items from compositional queries, and where the approach fails.

Introducing Per-Document Embedding Fine-Tuning

Talk
16. September 2026, 03:15 pm - 04:00 pm
Main Stage
We introduce per-document fine-tuning, a method that updates individual embeddings with new information while keeping the embedding model frozen. For example, applied to product search, it enriches product image embeddings with abstract attributes such as material and use case, improving retrieval without costly and risky model fine-tuning.

Beyond LLM Judges: Deep Evaluation for Conversational Search

Talk
15. September 2026, 10:00 am - 10:45 am
Main Stage
This talk presents a practical evaluation framework for an e-commerce conversational search assistant. In addition to a lightweight LLM-as-a-judge, we track additional metrics such as search term accuracy, filter precision and no-results rate. Learn how these signals surfaced concrete failure modes, guided our iterations and improved relevance.