Abstracts
Applications to speak. These go through review and a decision.
| Tags | Speaker | Ratings | Files | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Agents in Production: Field Notes from Cloudreach Labs | Pending | Approved | — | Abstract | — | A field report on running AI agents safely in production, covering guardrails, rollback, and observability lessons learned at Cloudreach Labs. | — | — | — | — | — | Sswyx+kms-0033-speaker2@ai.engineer | Sswyx+kms-0033-speaker2@ai.engineer | — | — | — | — | — | |
| Your AI Pair Programmer Is Lying to You: Verification Patterns That Scale | Pending | Approved | — | Abstract | — | Code generation is easy; trusting it is hard. This session covers verification patterns for AI-generated code — property tests, mutation coverage, snapshot judges, and CI gates — with data from 18 months of running them on a 200-engineer codebase. Includes what we stopped doing because it didn't catch anything. | — | — | — | — | — | PRPriya Raman | Sswyx+kms-0033-speaker@ai.engineer | — | — | — | — | — | |
| Docs That Answer Back: Retrieval-Grounded Documentation Sites | Pending | Approved | — | Abstract | — | A 10-minute tour of turning a static docs site into one that answers questions with citations, stays honest when it doesn't know, and costs under $50/month to run. Live demo, real failure cases, and a checklist you can apply to your own docs this week. | — | — | — | — | — | PRPriya Raman | Sswyx+kms-0033-speaker@ai.engineer | — | — | — | — | — | |
| Your AI Pair Programmer Is Lying to You: Verification Patterns That Scale | Pending | Approved | — | Abstract | — | Code generation is easy; trusting it is hard. This session covers verification patterns for AI-generated code — property tests, mutation coverage, snapshot judges, and CI gates — with data from 18 months of running them on a 200-engineer codebase. Includes what we stopped doing because it didn't catch anything. | — | — | — | — | — | Ssbek-speaker2@example.com | Ssbek-speaker2@example.com | — | — | — | — | — | |
| Claude - Event | Pending | In Review | — | Abstract | — | Test | — | — | — | — | — | Ffarishussain021@gmail.com | Ffarishussain021@gmail.com | — | — | — | — | — | |
| Closing Remarks | Pending | Approved | — | Abstract | AI-039 | Closing thoughts, thanks, and what we would like to see submitted next year. | Talk (30 min) | Advanced | English | Evaluation | Open SourceSecurity | SRSofia Rossi | SOSam Organiser | — | — | — | — | 0.9 | |
| Opening Keynote: Why Now | Pending | Approved | — | Abstract | AI-040 | Why this conference, why now, and what we hope you take away from the next two days. | Talk (30 min) | Beginner | English | Product | Open Source | WCWei Chen | SOSam Organiser | — | — | — | — | 1.2 | |
| Panel: The Ethics of Autonomy | Pending | Approved | — | Abstract | AI-038 | Autonomy raises questions the industry has been deferring. A panel on accountability, disclosure and the decisions we are quietly making on users' behalf. | Talk (30 min) | Intermediate | English | Infrastructure | CostSecurity | IKIdris Khan | SOSam Organiser | — | — | — | — | 0.2 | |
| Edge Inference on a Budget | Pending | Approved | — | Abstract | AI-009 | Running models close to users without a GPU budget. Quantisation choices, the memory ceiling on commodity edge hardware, and an honest account of the quality we traded away and where it turned out to matter. | Talk (30 min) | Advanced | English | Applied AI | Open SourceSecurity | GAGrace AdeyemiKMKofi Mensah | SOSam Organiser | — | — | 2.13 (4) | — | 1.2 | |
| Streaming UIs for Slow Models | Pending | Approved | — | Abstract | AI-008 | When the model takes eleven seconds, the interface is the product. Streaming, skeletons, optimistic rendering and the moment users decide something is broken — with the abandonment numbers that changed our minds. | Workshop (120 min) | Intermediate | English | Product | Cost | WCWei Chen | SOSam Organiser | — | — | 2.88 (4) · 1.75–4.25 | — | 2.8 | |
| Prompt Injection in the Wild | Pending | Approved | — | Abstract | AI-007 | A field report on prompt injection attempts against a public-facing assistant: what was tried, what worked, and which mitigations were theatre. Includes the payload that got past three layers of filtering. | Talk (30 min) | Beginner | English | Evaluation | AgentsCost | SRSofia Rossi | SOSam Organiser | — | — | 3.25 (3) | — | 1.5 | |
| Observability for Nondeterministic Systems | Pending | Approved | — | Abstract | AI-006 | Traditional observability assumes the same input gives the same output. We cover the tracing schema we settled on, how we sample when every request is unique, and how to alert on quality drift without drowning in false positives. | Talk (30 min) | Advanced | English | Infrastructure | RAG | IKIdris Khan | SOSam Organiser | — | — | 3.31 (4) · 2.25–4.5 | — | 2.6 | |
| Fine-tuning Is Not the Answer | Pending | Approved | — | Abstract | AI-005 | Fine-tuning is the first thing people reach for and usually the wrong one. We compare it against retrieval, prompt work and routing on the same three tasks, with the training costs and the maintenance burden included honestly. | Talk (30 min) | Intermediate | English | Applied AI | Open SourceRAG | NFNaomi FischerWCWei Chen | SOSam Organiser | — | — | 4.00 (4) | — | 0.6 | |
| The Cost Curve of Inference | Pending | Approved | — | Abstract | AI-004 | Inference costs do not scale the way finance expects. A breakdown of where our spend actually went across a year — prefill versus decode, cache hit economics, and the surprising fraction consumed by retries and abandoned streams. | Talk (30 min) | Beginner | English | Product | Open SourceSecurity | TBTomas Berg | SOSam Organiser | — | — | 2.63 (4) | — | 1.9 | |
| Vector Databases at a Billion Rows | Pending | Approved | — | Abstract | AI-002 | Everything is fast at ten million rows. We walk through what actually broke between one hundred million and a billion: index build times, memory-mapped segment churn, and the recall cliff nobody warns you about when you quantise too aggressively. | Talk (30 min) | Intermediate | English | Infrastructure | Agents | LMLuis Moreau | SOSam Organiser | — | — | 3.06 (4) | — | 1.2 | |
| Evaluating RAG: Beyond Vibes | Pending | Approved | — | Abstract | AI-003 | Most RAG evaluation is a demo and a feeling. We describe the offline harness we built, why we abandoned answer-similarity scoring, and how a small hand-labelled set of two hundred questions caught regressions our automated metrics happily approved. | Talk (30 min) | Advanced | English | Evaluation | CostSecurity | PRPriya Raman | SOSam Organiser | — | — | 3.00 (4) · 1–5 | — | 2.8 | |
| Shipping LLM Agents Without Losing Sleep | Pending | Approved | — | Abstract | AI-001 | We ran agents in production for eighteen months and most of what we believed at the start was wrong. This covers the retry semantics, the budget guards and the three incidents that shaped our current design, including the one that took a weekend to unpick. | Workshop (120 min) | Beginner | English | Applied AI | RAG | AOAda OkonkwoTBTomas Berg | SOSam Organiser | — | — | 3.31 (4) | — | 1.1 |
1–17 of 17
17 abstracts matching the current filters↑↓move⇧↑select rangespacetoggle⏎openescclear
1 / 1