read-only demo

Videos iRcX54EO5g8

Your agent is blindfolded — Johan Lajili, Poolside AI

index_state ready data_status ok

AI Engineer· published 2026-07-08· 0:09:57· en-US· indexed 2026-08-11 04:22

Open on YouTube

Scene timeline

  1. Shot 0, 0:00 to 0:05, 1 of 1 keyframes kept
  2. Shot 1, 0:05 to 0:09, 1 of 1 keyframes kept
  3. Shot 2, 0:09 to 0:14, 1 of 1 keyframes kept
  4. Shot 3, 0:14 to 0:36, 1 of 1 keyframes kept
  5. Shot 4, 0:36 to 0:39, 1 of 1 keyframes kept
  6. Shot 5, 0:39 to 1:05, 1 of 1 keyframes kept
  7. Shot 6, 1:05 to 1:30, 1 of 1 keyframes kept
  8. Shot 7, 1:30 to 1:56, 1 of 1 keyframes kept
  9. Shot 8, 1:56 to 2:22, 0 of 1 keyframes kept
  10. Shot 9, 2:22 to 2:47, 1 of 1 keyframes kept
  11. Shot 10, 2:47 to 3:16, 1 of 1 keyframes kept
  12. Shot 11, 3:16 to 3:44, 0 of 1 keyframes kept
  13. Shot 12, 3:44 to 4:11, 1 of 1 keyframes kept
  14. Shot 13, 4:11 to 4:37, 0 of 1 keyframes kept
  15. Shot 14, 4:37 to 5:04, 1 of 1 keyframes kept
  16. Shot 15, 5:04 to 5:30, 0 of 1 keyframes kept
  17. Shot 16, 5:30 to 5:57, 0 of 1 keyframes kept
  18. Shot 17, 5:57 to 6:24, 0 of 1 keyframes kept
  19. Shot 18, 6:24 to 6:58, 0 of 1 keyframes kept
  20. Shot 19, 6:58 to 7:33, 1 of 1 keyframes kept
  21. Shot 20, 7:33 to 7:37, 1 of 1 keyframes kept
  22. Shot 21, 7:37 to 8:02, 1 of 1 keyframes kept
  23. Shot 22, 8:02 to 8:27, 1 of 1 keyframes kept
  24. Shot 23, 8:27 to 8:52, 1 of 1 keyframes kept
  25. Shot 24, 8:52 to 9:17, 1 of 1 keyframes kept
  26. Shot 25, 9:17 to 9:42, 0 of 1 keyframes kept
  27. Shot 26, 9:42 to 9:57, 1 of 1 keyframes kept

27 shot(s).

keyframes kept every frame deduplicated

What was stored

cues
127
whisperx 127
chunks
18
from 127 cues
keyframes
19
kept of 27 captured
frames with text
19
255 lines read
chapters
0
from the source metadata
keyframe bytes
2.4 MB
word timings on 127 cues

Provenance

Each pipeline stage, its state and the model that produced it
stage state model started took
fetch done 2026-08-11 04:20 1m 11s
stt done 2026-08-11 04:21 10s
chunk done 2026-08-11 04:21 0s
text_embed done 2026-08-11 04:21 1s
keyframe done 2026-08-11 04:21 50s
ocr done 2026-08-11 04:22 6s
frame_embed done 2026-08-11 04:22 3s

Frames, and what the machine read

  • 0:02 #0 done2 line(s)

    shot 0·sharpness 668.7

    1. Al Engineer0.93
    2. EUROPE1.00
  • 0:06 #1 done2 line(s)

    shot 1·sharpness 819.1

    1. PRESENTING SPONSOR0.99
    2. Google DeepMind1.00
  • 0:11 #2 done3 line(s)

    shot 2·sharpness 939.2

    1. PLATINUM SPONSORS0.98
    2. # Braintrust0.95
    3. WorkOS OpenAI0.96
  • 0:17 #3 done5 line(s)

    shot 3·sharpness 412.0

    1. Your Agent Is1.00
    2. Blindfolded1.00
    3. Johan Lajill0.95
    4. plside0.97
    5. ineer1.00
  • 0:37 #4 done11 line(s)

    shot 4·sharpness 1405.2

    1. AIE1.00
    2. 1.00
    3. 1.00
    4. 1.00
    5. 0.99
    6. Your Agent Is0.98
    7. Blindfolded1.00
    8. How giving it (good) eyes multiplies performance and trust0.99
    9. Johan Lajili1.00
    10. Member of Engineering (Full Stack) - Poolside - London0.99
    11. Google DeepMind1.00
  • 0:57 #5 done20 line(s)

    shot 5·sharpness 2626.4

    1. r/ChatGPTCoding Posted by u/mass_shipped_vibes0.99
    2. I haven't written a line of code in 3 months. Cursor + Claude does everything.1.00
    3. I just review and ship. This is the future and I'm never going back.0.98
    4. *★0.60
    5. 847 comments0.99
    6. AIE1.00
    7. 1.00
    8. 1.00
    9. 1.00
    10. 0.99
    11. VS.0.98
    12. r/ExperiencedDevsPosted by u/legacy_code_survivor0.99
    13. Genuinely curious what codebase these people are working on. I tried Claude0.99
    14. on our monorepo and it hallucinated half the imports, broke 3 services, then0.99
    15. told me everything was passing. Al coding tools are a mass delusion for1.00
    16. anything beyond todo apps.1.00
    17. 1.2k comments1.00
    18. podside0.95
    19. Engineering the future of Al0.99
    20. gineer0.92
  • 1:18 #6 done20 line(s)

    shot 6·sharpness 2725.1

    1. r/ChatGPTCoding· Posted by u/mass_shipped_vibes0.98
    2. I haven't written a line of code in 3 months. Cursor + Claude does everything.0.99
    3. I just review and ship. This is the future and I'm never going back.0.99
    4. *★0.69
    5. 847 comments0.98
    6. AIE1.00
    7. 1.00
    8. 1.00
    9. 1.00
    10. 0.99
    11. VS.1.00
    12. r/ExperiencedDevsPosted by u/legacy_code_survivor1.00
    13. Genuinely curious what codebase these people are working on. I tried Claude0.99
    14. on our monorepo and it hallucinated half the imports, broke 3 services, then0.99
    15. told me everything was passing. Al coding tools are a mass delusion for1.00
    16. anything beyond todo apps.1.00
    17. 1.2k comments0.99
    18. Braintrust1.00
    19. WorkOS OpenAI0.95
    20. IEngineer0.96
  • 1:51 #7 done21 line(s)

    shot 7·sharpness 2610.1

    1. r/ChatGPTCoding1.00
    2. Posted by u/mass_shipped_vibes0.97
    3. I haven't written a line of code in 3 months. Cursor + Claude does everything.1.00
    4. I just review and ship. This is the future and I'm never going back.0.99
    5. ***0.57
    6. 847 comments0.99
    7. AIE1.00
    8. 1.00
    9. 1.00
    10. 1.00
    11. 0.99
    12. VS.0.98
    13. r/ExperiencedDevsPosted by u/legacy_code_survivor1.00
    14. Genuinely curious what codebase these people are working on. I tried Claude0.99
    15. on our monorepo and it hallucinated half the imports, broke 3 services, then0.99
    16. told me everything was passing. Al coding tools are a mass delusion for1.00
    17. anything beyond todo apps.1.00
    18. 1.2k comments1.00
    19. AlEngineer0.96
    20. EUROPE1.00
    21. AlEngineer1.00
  • 2:04 #8 skipped

    shot 8·duplicate of #7

  • 2:42 #9 done20 line(s)

    shot 9·sharpness 2604.2

    1. r/ChatGPTCoding Posted by u/mass_shipped_vibes0.98
    2. I haven't written a line of code in 3 months. Cursor + Claude does everything.1.00
    3. I just review and ship. This is the future and I'm never going back.0.98
    4. *★0.63
    5. 1.00
    6. 847 comments0.98
    7. AIE1.00
    8. 1.00
    9. 1.00
    10. 1.00
    11. 0.99
    12. VS.0.83
    13. r/ExperiencedDevs· Posted by u/legacy_code_survivor0.97
    14. Genuinely curious what codebase these people are working on. I tried Claude0.99
    15. on our monorepo and it hallucinated half the imports, broke 3 services, then1.00
    16. told me everything was passing. Al coding tools are a mass delusion for1.00
    17. anything beyond todo apps.1.00
    18. 1.2k comments1.00
    19. Engineering the future of Al0.99
    20. AlEngineer0.98
  • 3:10 #10 done11 line(s)

    shot 10·sharpness 1293.5

    1. "I've implemented the new OAuth flow, it's all working”0.99
    2. To the best of the capabilities you have given me, this sounds1.00
    3. AIE1.00
    4. like it would work.1.00
    5. 1.00
    6. 1.00
    7. 1.00
    8. 0.99
    9. poolside1.00
    10. Engineering the future of Al0.99
    11. Engineer1.00
  • 3:33 #11 skipped

    shot 11·duplicate of #10

  • 3:52 #12 done20 line(s)

    shot 12·sharpness 2280.0

    1. Our CLl at Poolside: Spoolside0.98
    2. Giving the agent eyes into the running app1.00
    3. >_ Extract text & element snapshots0.98
    4. Screenshots — page or component1.00
    5. Interact with components by ref0.98
    6. </> Trace DOM → Svelte components0.95
    7. AIE1.00
    8. Restart services locally1.00
    9. Extract logs — backend, helper, FE0.98
    10. 1.00
    11. 1.00
    12. 1.00
    13. 1.00
    14. High-level: "Open settings menu"0.99
    15. "Change feature flag", "Seed DB"0.98
    16. But the point is not our tools, it's yours1.00
    17. Make tools to interact with your product specific to your product.1.00
    18. AlEngineer0.97
    19. EUROPE1.00
    20. AlEngineer0.99
  • 4:32 #13 skipped

    shot 13·duplicate of #12

  • 4:40 #14 done21 line(s)

    shot 14·sharpness 2267.8

    1. Our CLI at Poolside: Spoolside0.98
    2. Giving the agent eyes into the running app1.00
    3. >_ Extract text & element snapshots0.98
    4. Screenshots — page or component1.00
    5. ***0.70
    6. 1.00
    7. Interact with components by ref0.98
    8. </> Trace DOM → Svelte components0.97
    9. AIE1.00
    10. 1.00
    11. Restart services locally1.00
    12. Extract logs — backend, helper, FE0.99
    13. 1.00
    14. 1.00
    15. 1.00
    16. High-level: "Open settings menu"1.00
    17. "Change feature flag", "Seed DB"1.00
    18. But the point is not our tools, it's yours1.00
    19. Make tools to interact with your product specific to your product.1.00
    20. Engineering the future of Al1.00
    21. AlEngineer0.99
  • 5:12 #15 skipped

    shot 15·duplicate of #14

  • 5:54 #16 skipped

    shot 16·duplicate of #12

  • 6:05 #17 skipped

    shot 17·duplicate of #12

  • 6:34 #18 skipped

    shot 18·duplicate of #10

  • 7:22 #19 done17 line(s)

    shot 19·sharpness 1377.8

    1. 20251.00
    2. 20261.00
    3. AIE1.00
    4. Product1.00
    5. aiX1.00
    6. 1.00
    7. Engineer1.00
    8. Engineer1.00
    9. 1.00
    10. 1.00
    11. 1.00
    12. Building the product < Making sure the Al can build the product1.00
    13. Thanks.1.00
    14. polside0.96
    15. WorkOs OpenAI0.95
    16. Braintrust1.00
    17. AlEngineer0.98
  • 7:34 #20 done35 line(s)

    shot 20·sharpness 2166.2

    1. 0.53
    2. Slide0.97
    3. Our CLI at Poolside: Spoolside0.98
    4. DEFAULT1.00
    5. Appearance1.00
    6. Title0.94
    7. Giving the agent eyes into the running app1.00
    8. Body1.00
    9. Slide Number0.96
    10. >_ Extract text & element snapshots0.97
    11. Screenshots — page or component0.99
    12. Background1.00
    13. Standard1.00
    14. Dynamic0.99
    15. Current Fill0.98
    16. AIE1.00
    17. 1.00
    18. 1.00
    19. Interact with components by ref1.00
    20. <>0.95
    21. Trace DOM → Svelte components0.98
    22. Colour Fll0.87
    23. 1.00
    24. Restart services locally1.00
    25. Extract logs — backend, helper, FE1.00
    26. 1.00
    27. 1.00
    28. 1.00
    29. High-level: "Open settings menu"0.98
    30. "Change feature flag", "Seed DB"0.99
    31. But the point is not our tools, it's yours0.99
    32. Make tools to interact with your product specific to your product.0.99
    33. AlEngineer0.97
    34. EUROPE1.00
    35. AlEngineer0.97
  • 7:49 #21 done9 line(s)

    shot 21·sharpness 438.0

    1. 20261.00
    2. Product1.00
    3. aiX0.99
    4. Engineer1.00
    5. Engineer1.00
    6. Thanks.1.00
    7. olside0.92
    8. AlEngineer0.98
    9. EUROPE0.98
  • 8:17 #22 done9 line(s)

    shot 22·sharpness 438.7

    1. 20261.00
    2. Product1.00
    3. aiX1.00
    4. Engineer1.00
    5. Engineer1.00
    6. Thanks.1.00
    7. poolside0.99
    8. Engineer1.00
    9. UROPE1.00
  • 8:49 #23 done9 line(s)

    shot 23·sharpness 435.4

    1. 20261.00
    2. Product1.00
    3. aiX0.99
    4. Engineer1.00
    5. Engineer1.00
    6. Thanks.1.00
    7. pocside0.98
    8. Engineer1.00
    9. ROPE0.97

Transcript

127 cues· 1,602 words· 8,399 chars

  1. 0:15 Hi, everyone.
  2. 0:16 So I'm Johan from Poolside.
  3. 0:19 If you haven't heard of us, we are one of the handful of companies that are making their own foundational model from scratch, their own LLM and coding agents.
  4. 0:31 Check us out if you're not aware.
  5. 0:33 Some cool stuff, and you should hear more soon.
  6. 0:37 But what I want to talk to you about today is this.
  7. 0:44 people seeing AI and using AI and getting vastly different experiences.
  8. 0:48 If you're on Reddit, if you're on Twitter, you're going to see people that say, oh, yeah, I'm never touching code anymore.
  9. 0:54 The AI is doing everything for me.
  10. 0:56 It's fantastic.
  11. 0:58 And others that say, what are you talking about?
  12. 0:59 I'm trying it in my production app.
  13. 1:02 It produces absolute garbage.
  14. 1:05 What are you guys working on to do apps?
  15. 1:07 What are you doing?
  16. 1:08 And there is multiple ways to try
  17. 1:11 to understand what's going on there.
  18. 1:13 One way is to say, oh, this guy is a shield from OpenAI trying to sell you AI.
  19. 1:18 Or this one is an NT that doesn't care about anything and is just lying.
  20. 1:23 He didn't even try it.
  21. 1:25 Another is to say,
  22. 1:28 Well, maybe someone is working on a Greenfield app, and that's nice and easy.
  23. 1:32 Whereas someone else is working on Brandfield on a legacy application, that's complicated, and agents are not there yet.
  24. 1:40 But personally, I think that doesn't quite hold up.
  25. 1:44 We've seen people using AI in legacy applications with good success.
  26. 1:48 I have myself.
  27. 1:50 At least on my own experience, I know that can work.
  28. 1:54 So what's the difference?
  29. 1:54 What's the difference really between Brownfield and Greenfield?
  30. 2:00 The difference is that with Greenfield, the agent's intuition is correct.
  31. 2:06 the agent's writing the code and expect, you know, if I write the components here, if I write the service, it's going to work.
  32. 2:12 I think that's going to be fine.
  33. 2:14 And he's right, because it has very good intuition.
  34. 2:17 On brownfield, however, there be dragons.
  35. 2:21 You're going to have things that the agent is not expecting.
  36. 2:25 Maybe, you know, like dead ends, code that's not used anymore, things that he's not aware of in different parts of the code that he hasn't even looked at.
  37. 2:33 And that's where the big difference between those two is the feedback loop.
  38. 2:37 And everybody sort of like somewhat talked about it in the background of the talks we've seen over the past three days.
  39. 2:44 But I think that's the difference with getting these results.
  40. 2:48 So you've got your agent that says, yeah, I've implemented the new OAuth flow and it's all working perfectly.
  41. 2:55 What the agent means really is
  42. 2:58 Well, to the best of my capabilities, to the best of what you have given me, that sounds like it should work.
  43. 3:05 Maybe the agent was able to verify its work.
  44. 3:08 Maybe it wasn't.
  45. 3:09 But as far as it knows, it's working.
  46. 3:12 If you're a skeptic of AI, you're going to see
  47. 3:15 that first quote sees that it's indeed not working and just says, you know, I'm a liar, I'm incompetent or like you cannot trust the AI.
  48. 3:26 And that's, I think, where this cleavage in between those two type of users, like the first category is going to see some things.
  49. 3:32 Oh, yeah, actually, you know what?
  50. 3:33 It's not working, agent.

Open at this second