ppll

Related posts

Eugene Yan Bluesky

If you were building a Q&A feature (or chatbot) based on very long documents (like books), what evals would you focus on?

chatbotevalsdocuments

The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.

Sources

Top people

18 September

Hamel Husainx.com

  • Related

Alex Tabarrokx.com

  • Related

17 September

Steven Johnsonx.com

  • GoogleRelated

16 September

15 September

Hamel Husainx.com
  • hamel.dev

    LLMsRelated

14 September

Karri Saarinenx.com

  • Related

11 September

Jakob NielsenNewsletter

9 September

Cassidy Williamsx.com

  • fandf.co

    Related

6 September

Karri Saarinenx.com
  • Related

5 September

4 September

Hamel Husainx.com
  • Related

  • Related

2 September

Teresa TorresBlog

1 September

Ajeya CotraKorrents

31 August

Hamel Husainx.com
  • Related

Joe LiemandtKorrents

From one piece How to Accelerate Learning & Improve Education | Joe Liemandt 2 beliefs, in the piece's order there

28 August

Sriram Krishnanx.com

  • OpenAIHuggingFaceRelated

26 August

Arvid Kahlx.com

  • Related

21 August

Karri Saarinenx.com
  • Related

17 August

Federico ViticciBlog

  • How Should the Siri App Be Graded?

    Related

16 August

12 August

Hamel HusainSite

10 August

Patrick McKenziex.com

  • LLMsRelated

27 July

Boris ChernyKorrents

25 July

Dan LuuMastodon

  • Related

24 July

Dan Luu
  • x.com

    Related

  • Bluesky

    danluu.com

    Related

  • Mastodon

    danluu.com

    Related

23 July

Ethan MollickKorrents

21 July

Adam MastroianniRecommends

  • Recommendsrcmnd.app

    Chatting Alone

    Their words

    Martsinovich argues, via code snippets, that there is no such thing as a "conversation" with a chatbot. The AI is born anew every time it speaks, and it simply reads the dialogue so far and then tries to write the next line:

    experimental-history.com

19 July

15 July

Dex HorthyKorrents

1 July

Eugene Yanx.com

  • AnthropicRelated

29 June

Hamel HusainSite
  • “It’s Hard to Eval” Is a Product Smell

    Related

26 June

21 June

Eugene YanSite

8 June

Dan Luu
  • x.com

    Related

  • Mastodon

    patreon.com

    Related

4 June

Bruce Schneierx.com

Alex ImasKorrents

24 May

Dan ShipperKorrents

From one piece AI predictions: Job markets, Codex beats Claude, and the death of org charts | Dan Shipper 2 beliefs, in the piece's order there

13 May

Eugene Yanx.com
  • xbow.com

    Related

10 May

Eric RiesKorrents

From one piece How Anthropic, Costco, and Patagonia all build incorruptible companies | Eric Ries 2 beliefs, in the piece's order there

10 April

2 April

Andrej Karpathyx.com

  • LLMsRelated

29 March

Chip HuyenKorrents

20 March

Andrej KarpathyKorrents

4 March

2 March

Hamel HusainSite
  • Evals Skills for Coding Agents

    coding agentsRelated

13 February

Dario AmodeiKorrents

5 February

Mitchell HashimotoKorrents

23 January

Adam WigginsSite

12 December 2025

Marc Littlemore

  • Mastodon

    AnthropicRelated

  • Bluesky

    AnthropicRelated

25 November 2025

Ilya SutskeverKorrents

Eleanor JanegaKorrents

23 November 2025

Andrew BatsonBlog

  • China bows to de-industrialization

    ChinaRelated

Eugene YanSite

27 October 2025

Jasmine SunNewsletter

1 October 2025

Hamel HusainSite
  • Selecting The Right AI Evals Tool

    Related

19 September 2025

Norman OhlerKorrents

2 September 2025

Benedict EvansKorrents

29 August 2025

5 August 2025

25 June 2025

Eugene YanBluesky
  • eugeneyan.com

    LLMsbenchmarksRelated

22 June 2025

28 May 2025

Hamel HusainSite
  • AI Evals: Everything You Need to Know

    Related

12 May 2025

30 April 2025

Eugene YanBluesky
  • maven.com

    Related

23 April 2025

Eugene YanBluesky
  • eugeneyan.com

    LLMsRelated

16 April 2025

Eugene YanBluesky
  • Related

26 March 2025

25 March 2025

Julie ZhuoKorrents

3 February 2025

Nathan LambertKorrents

29 October 2024

Hamel HusainRecommends
  • Lovedrcmnd.app

    Hex

    Their words

    You might be skeptical of using synthetic data. After all, it’s not real data, so how can it be a good proxy? In my experience, it works surprisingly well. Some of my favorite AI products, like Hex use synthetic data to power their evals

    hamel.dev

3 July 2024

Nelson ElhageBlog

21 June 2024

Simon WillisonRecommends

  • Their words

    Your AI Product Needs Evals by Hamel Husain remains my favourite piece of writing on how to go about putting these together.

    simonwillison.net

19 June 2024

Aravind SrinivasKorrents

3 June 2024

What is a korrent?

A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.

Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.

Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.

About the English under a post

Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.

The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.

Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.