One topic. Every takeSeek and you shall find

Reading up on Weco’s AIDE2

1 deep · digging since sep 10

  • www.chrismdp.com favicon
    Prompt Evals Alone Are Useless

    Prompt evaluations by themselves cannot ensure LLM app quality; you must test the entire system—harness, memory, and UI—through code-driven scenarios.

read later

For all the tabs you promised to read.
Save to read. Read to clear.

Close tabs. Keep links.