Transcript: 8G 1 3Io4Zq
Source Video
Local Cache
raw/sources/youtube-transcripts/8G_1-3IO4ZQ.txt- 3,420 words
Transcript
Hi everyone. Uh my name is Praalpa. I'm the founder of Atlan. Um and uh today I'm going to talk about this thing where context is having its moment. Uh and so my goal today is to talk about like WTF is the context layer. Um, just before I start, and I promise this is the last time. Um, I don't know if the clicker is working. Atlin, we it's it's working. Yeah. Thank you. Um the problem we solve is we say AI doesn't know your business. We fix that. We work with an incredible group of companies around the world ranging from GitLab and Zoom and Discord and Affirm to large enterprises like Mastercard and General Motors. Um and about a year ago, uh my co-founder and I went on stage and we said, uh at the dawn of the internet era, Bill Gates had written this very famous blog post and it said content is king. Um and as we or the dawn of the agentic era, context will be king. Um since then it feels like 2026 is the year of context. context graphs anyone um uh you know every every two days you see some version of context uh popping up and so what is going on um I believe the answer to this kind of is in this reality distortion field that we live in uh I live here in the Bay Area every day or two I have conversations with people which kind of go like how far are we from AGI and we have a debate and we're like well one year three years so on uh There is no doubt that the models are getting exponentially smarter by the day. Uh two years ago they couldn't pass the bar. Today if they were to take the bar it was they're the top 1% of test scorers. On the other hand they're not exponentially more useful by any benchmark. Uh one out of five you know AI use cases actually make it to production. U you know 56% of CEOs say that there's zero financial benefit from AI today. So what's going on? I believe hidden in plain sight is actually um how performance is measured in the human world. Uh cognitive intelligence doesn't really determine real world effectiveness. Uh in fact only 10% of job performance variance is explained by IQ. Like just think about it. Would you say your smartest um you know teammate who scored the highest on the SATs is also your best teammate or would you say no it's the person who works the most and takes the most feedback and learns the fastest in the real world we care about performance and performance is outcomes that you deliver in the real world and performance is a function of two things it's a function of intelligence which is cognitive horsepower that's what the model benchmarks measure every day But it's also a function of context. This is what they say in the human world as learning on the job, right? Knowledge and skills and expertise that you learn over time. And in the last decade, uh we have compounded on one of those parameters. Uh intelligence has thousandxed in the last decade. Just in the last 6 months, we have 2xed on that axis. On the other hand, context, the situated knowledge of your business, that's barely moved. We've moved some data to the cloud uh but that's about it. It's otherwise logged in dashboards and Slack threads and uh the head of that analyst who might be leaving next week. Um and so the question ahead of us and I really believe this is the next frontier is how do we help AI build context about our business? Um, and every time I'm faced with a question about how do we help AI do this, I always like to go back and understand how did we help humans do this? Uh, so I'm going to take you into the life of, you know, a u exemplar employee Maya. Uh, let's say she's a data analyst at Mech Context Burgers because I thought I was going to be creative and I'm not very creative. Um, and you know, let's say she's that analyst that everybody, you know, pings in your company. Uh, right? She's the person that everybody sends a message to every morning when they're trying to solve a problem. So, let's say this morning, uh, there's a franchisee owner who sends her a message and says, "Why is my drive-thru time up this week? Why is this metric up this week?" Sounds like a really simple question. Um, but it's actually a really complicated question to ask. Just to answer this one very simple question, Maya first needs to know uh what is drive-through time uh and who's asking? Is it finance or is it you know my ops team? And it might mean different things. Uh but not just that, what does this week mean? Is the cutoff period Monday to Sunday? Is it Pacific time? Is it Eastern time? Uh that's knowledge. Like that's facts. That's the map of the business. Um but not just that. Uh there's expertise uh right there's um you know a diagnostic playbook. What what does a great analyst do? They know that you know quarter 3 is a season seasonal quarter because of weather patterns and they know to go check if the reason there's a spike is because of seasonality. They also know that the company launched a product uh just that previous quarter and so they know to check if that's why the root cause analysis failed. Uh this is expertise and skills that people pick up over time as they learn on the job. Uh and then there's norms, right? Um there's, you know, persona scoping. Who's asking the question? How do I answer this question? Um and Maya, she's one of those like cool people. She nails it. She sends an answer not just with the answer, but with the why and the root cause, and she finds the reason for it. How did Maya learn to do this? Um she just joined the company a year ago. Um first Maya you know has for like she joined and she got some training like all of us do but that's not where any of us learn right in our companies. How do we learn? We learn because you shadow like the best teammate and then you see why they're doing something and then you learn from that and then you make a mistake. Who here has learned more from a mistake than anything else? Right? You make a mistake and then you learn. uh you your manager gives you feedback and you learn not to do that again. You deal with an edge case and then you learn from that. That's how all of us humans learn at work. And so then the question is how do you help build the agent Maya? Uh and now I want to walk you through our experiments and learnings as we've built this at Atlan um era one and this was roughly about 18 months ago now. Um we uh started on the the track of bootstrapping agents. Um and the way we went about it was and we started this with our customer experience team. Uh and we did this jobs to be done analysis map, right? And so we said, hey, if you are someone on our customer experience team, what are all the things that you do on a day-to-day basis? And then we made some hypothesis. We we said, you know, for example, one part of the job is documentation and meeting prep. Uh we said well AI could probably do that job pretty well. Uh and so we build a scaling factor. So on the other hand relationship management is something that our customer experience team does and we said hm that doesn't sound like something AI is going to be able to do anytime soon. And so we built a scaling factor and then we basically started bootstrapping these individual agents that were like built for that specific topic. Uh our team got creative. So we had Hermione who is our health intelligence lead and then we had you know money penny who was our financial risk analyst and we just made that particular agent really good at doing that one thing. Um and that worked for some time um but then we realized there were some challenges with this approach. The first context engineering uh we got to the point by middle of last year where building an agent was really easy took like 5 minutes. uh but giving it the business context that it took to actually get it to be accurate took forever. Uh quality of the agent often dependent uh on the quality of context engineering and that led to a lot of weird lost trust cases with our stakeholders. Um then as we started taking this into production we started seeing that these agents basically were kind of like living on their own island. Uh now imagine for example if you're in a human team and your marketing changes positioning on your you know and then they come to the town hall and they tell you that they changed positioning and so then you know the SDR on your team or your sales development rep they know that they should use that new positioning. This is like the infrastructure that we've built for humans inside our organizations. Agents didn't have them. So our marketing team had these agents and they started making changes to that and then our SDR agent on our website was still pitching the old version. Uh we had no idea how any of these things were even connected. So we didn't even know how to like run this as as a team of agents. Uh when an agent gets something wrong, this is hard. Uh it was really hard to like trace back what happened. Was it the model? Was it the agent? Was it the context? Like where how do we even go back and fix this? Um and over time we started dealing with uh context sprawl. Uh we had the the the hard part about this was agents all had their own memory systems to a certain extent. So they were learning they were all learning separately and they were learning differently. Uh it became very very difficult very quickly to say okay what does the single version of truth here look like? Um and then over time we actually went through in the last 12 months we've gone through cycles of at the agentic layer about 12 months ago we were using one of these no code type builders uh called relevance we went from there into Google ADK then we tried glean uh start of this year we moved to cloud code now we are kind of like 50/50 claude and codeex um and every single time as these changes happened uh our context got trapped in each of these individuals systems. Um, so started this year as general purpose agents started to become a thing, we said, what if there was a different approach with general purpose agents. Um, again going back to the human world, well Maya, she's not an individual star. She's part of a team, right? And you know, you talk about these dream teams like Maya and someone who runs customer support and someone who launches ads. These people work really well together. And often these dream teams are built on shared context, right? Uh they have a shared language. Uh they have a shared picture of what's true today. They have shared playbooks. Uh they have shared norms, who's allowed to make what decision. Uh and then they learn together. I think this is the most important part of it. They have compounding learning loops of what good looks like. uh and they have shared memory that you know oh we launched this thing last quarter and it like was terrible and we're not going to make that mistake again right and so we said is there a way to bring that into the way we think about AI in our companies and so the mental model we started working on was we said okay we have these teams of humans and they're across the board and can these people essentially start building domain skills so each of them is responsible for a certain set skills. All of this goes into this common one place which is this one company brain of sorts, right? I like to think of this as the context layer. Uh and then this has a bunch of retrieval mechanisms which then talks to the general purpose agent across the ecosystem. So then we started an experiment. Uh this is some version of what our marketing team ended up building. So you'll see on the left those are all the systems that our marketing team uses. So data systems, our social and community platforms, our ad platforms, our analytics platforms. Um and then you'll see this agent block. Uh we built this very specifically for um having openness. So we had claw code and co-work. We also had our own claw that we deployed which has you know essentially talks in our slack channels. Um and then we used some external products like qualified and artisan. Uh in the middle is kind of this context layer that our team started building. So think of it as our best SEO person was building their SEO skill. Uh our best competitive intel person was building the best competitive intel skill and that kind of became this common repo that we were building into and pulling out from. This sort of became our living brain. Over time, we realized there were some things that we needed in this brain, right? Uh we realized we needed a data graph like if for example our autonomous ads agent, we realized it needs to do analysis on a daily basis. So like which table should I go pull from? Uh we needed a library of skills. We also needed some other things, semantics, metrics, what is ARR, how do you measure that? Uh what is a qualified lead in our company? uh and or structure entities things like that. Over the last 6 months, we ended up creating about 300 skills and 40 agents in this team. Uh which has been incredible. Uh but then with this approach too, we realized that there were some challenges. We realized that context kind of needs to be managed like code. Um so some challenges, let's pick skills. uh dependency management became really complicated. So for example, we have this comparative intelligence skill and it learns from the market on what's changing in the market and it improves. Um it feeds our category positioning skill which then feeds our sales battle card skill. Uh now each of these skills is learning and evolving. Uh but every time they learn and evolve it breaks something downstream. uh and these skills very quickly start getting outdated and start drifting. Uh who owns skill quality became another thing like who eventually owns the quality of this security and governance was a nightmare. Uh we had secrets hardcoded in ENV files. Uh it was people were downloading these public skill repos. This the whole thing was like a nightmare. Um and then I talked about context portability across all these multi-agent systems. I started this talk by saying WTF is a context layer. Uh these are the problems that a context layer is meant to solve. Um the question I like to ask is what does the GitHub for context look like? Um few thoughts. Uh company context needs life cycle management, collaboration and versioning. Uh just like code does. uh you know there's questions like what's local context what's global context how do I keep this updated so on uh some thoughts in this can skills have a profile just like code does uh can that have a self-learning learning loop that's baked into it uh what does quality management look like can you have security and postures posture management associated with that that's really like the first step uh I see this as like having something that has built-in versioning and quality and dependency management. So you should be able to say, "Hey, this thing impacts all these other things. This is the approver. This is the maintainer. These are the contributors. How do you build like kind of human plus AI workspaces that these that these skills uh are managed via? Second thing, every AI interaction creates more context and harnessing this uh is gold. uh uh there's been I know a lot of talks about self-improving loops uh we have found that with traces deploying a specific harness that actually is specialized in being able to go and reverse construct from that. So think of it as AI that's reading through all your traces and almost brings it back to your maintainer loop and says approve reject approve reject improve this over time. Uh that's the compounding learning loop. And the third often a lot of people ask me this question which is like how do I start because my business is like really disperate and I have all these like 60 systems and how do I even start? One of the biggest learnings we've had is context is hidden in these in business systems. Uh and across this context quality can really compound. So for example, if you're able to connect your Salesforce and your HubSpot to your data warehouse to your application layer and then you're able to reverse construct how these things are actually connected one to another, context today gets lost in every one of those hops. But if you can reverse construct that and then deploy AI on top of it, we've seen incredible accuracy in being able to reverse construct the first version of your company brain. So I'll end with this. Uh the way I think about a context layer is it's a system that turns knowledge and expertise and norms that we talked about that Maya knows into a machine usable context for AI systems. Um at a very high level the way I like to think of it is it looks like this. Uh it continually is mining context from your business systems. It's feeding this in to that one company brain. It's harnessing this in skills and context development life cycles as your teams go and deploy these agents. And then it has a bunch of ways you can retrieve it. So MCP, SQL, vector retrieval, hybrid assembly, all these different ways that you retrieve it and pull back from traces and build this compounding learning loop. Today we're largely building agents by hard- coding context. The scale of this problem I truly believe is unhived because with scale this can become really unsustainable. Uh and a little dangerous like all of us know this this old joke which is if you ask sales and finance the revenue number you're going to get two different numbers. uh we're fast approaching a moment of starting to deploy autonomous systems where the same thing is starting to happen. So I'll end with one last thing. I started this presentation by saying context is king. Um I'd like to end it by saying context is also IP. Something I think a lot about is in a world where you and your competitor have access to the same models and the same intelligence, what differentiates a company? What differentiates a customer support agent at American Express versus Amazon? Uh that's how you do business. That's what makes your company special. uh context is how we take and encode our culture and our norms into something that we will be proud of as we build autonomous frontier firms. Um and that's all I had. Um you can find me at proalpa on Twitter um or write to me. We are actively working with folks on the frontier ongoing and shipping and building company brands. Um so if you'd like to talk to us, feel free to reach out. Thank you.