ppll

Related posts

Salvatore Sanfilippo GitHub

qwen-asr — C inference for Qwen3-ASR 0.6b and 1.7b transcriptions models

qwenqwen3inference

The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.

Sources

Top people

20 September

Nathan LambertBluesky

  • ChinaRelated

19 September

Nathan LambertBluesky
  • interconnects.ai

    Related

Guillermo Rauchx.com

  • OpenAIRelated

18 September

Dax Raadx.com

  • Related

  • Related

17 September

Tomasz TunguzKorrents

15 September

Sebastian Raschkax.com

  • LLMsOpenAIRelated

Paul GrahamKorrents

From one piece @paulg on X 2 beliefs · x.com

13 September

Ismail Ghalloux.com

  • llmgateway.io

    LLMsRelated

Andrew Chenx.com

  • LLMsAnthropicRelated

12 September

Xe IasoSite

11 September

Daniel Lemirex.com

  • LLMsRelated

10 September

Kun Chenx.com

  • AnthropicRelated

8 September

Byrne HobartNewsletter

6 September

Sebastian Raschkax.com
  • LLMsRelated

3 September

Tomasz Tunguzx.com
  • Related

Dax RaadKorrents

23 August

20 August

Kenneth Reitzx.com

  • kennethreitz.org

    Related

11 August

Sebastian Raschkax.com
  • LLMsRelated

30 July

Sebastian Raschkax.com
  • AnthropicRelated

Jeff DeanKorrents

From one piece Jeff Dean: The 1% Rule for Building in AI 3 beliefs, in the piece's order there

22 July

18 July

Sebastian RaschkaKorrents

30 June

29 June

Brad DeLongRecommends

  • Recommendsrcmnd.app

    llama3.2:3b

    Their words

    Right now: llama3.2:3b appears to be the model for: is this email urgent?qwen3:8b appears to be the model for: summarize this 5000-word article. llama3.3:70b(q8) appears to be the model for: let’s write or debug some computer code.

    braddelong.substack.com

27 June

Sebastian RaschkaRecommends

15 June

Sunil PaiBlog

Vicki BoykisRecommends

  • Usesrcmnd.app

    pi

    Their words

    For my local setup, I'm currently using Pi as the agent harness and LM Studio as the inference server

    vickiboykis.com

  • Usesrcmnd.app

    LM Studio

    Their words

    For my local setup, I'm currently using Pi as the agent harness and LM Studio as the inference server

    vickiboykis.com

9 June

4 June

Wes BosBluesky

  • blog.cloudflare.com

    Related

27 May

Dax RaadKorrents

From one piece Building OpenCode with Dax Raad 3 beliefs, in the piece's order there

25 May

  • Distributing LLM inference in DwarfStar

    LLMsRelated

22 May

Dan Luux.com

  • Related

13 May

Anders HejlsbergKorrents

2 April

Vitalik ButerinRecommends

  • Usesrcmnd.app

    llama-server

    Their words

    As it turned out, ollama was not able to fit Qwen3.5:35B onto my GPU, but llama-server could. Hence, from that day forward, I resolved to cease being a cave-dwelling noob, and use llama-server (via llama-swap to make model swapping easier).

    vitalik.eth.limo

  • Usesrcmnd.app

    Qwen3.5:35B

    Their words

    I have been using the Qwen3.5:35B model and have tried it on each of these, and I also tried the one-step-larger 122B.

    vitalik.eth.limo

23 March

13 March

Dylan PatelKorrents

From one piece Dylan Patel — The single biggest bottleneck to scaling AI compute 2 beliefs, in the piece's order there

15 February

13 February

Dario AmodeiKorrents

10 February

Casey HandmerKorrents

24 January

30 December 2025

28 December 2025

Peter SteinbergerBlog

4 December 2025

Anil SethPapers

17 September 2025

Eugene YanBluesky

  • eugeneyan.com

    Related

6 September 2025

15 August 2025

Casey HandmerKorrents

9 August 2025

31 July 2025

Peter SteinbergerBlog

3 February 2025

Dylan PatelKorrents

30 January 2025

Armin RonacherRecommends

  • Likedrcmnd.app

    MacBook Pro M1 Max

    Their words

    This allows me to run models locally on my MacBook Pro M1 Max. With the 64GB of RAM it has, it’s a pretty potent machine for basic inference despite it being three years old.

    lucumr.pocoo.org

31 December 2024

Simon WillisonRecommends

6 August 2024

19 June 2024

Aravind SrinivasKorrents

10 January 2023

Lilian WengBlog

  • Large Transformer Model Inference Optimization

    Related

What is a korrent?

A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.

Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.

Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.

About the English under a post

Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.

The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.

Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.