qm – Multiplayer agent harness for work
499 by tosh | 106 comments on Hacker News.
Friday, July 31, 2026
DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
494 by theanonymousone | 272 comments on Hacker News.
https://ift.tt/gk4smlB
494 by theanonymousone | 272 comments on Hacker News.
https://ift.tt/gk4smlB
Thursday, July 30, 2026
Wednesday, July 29, 2026
Tuesday, July 28, 2026
Netflix employee fired for sharing personal details in retreat trust exercise
Netflix employee fired for sharing personal details in retreat trust exercise
374 by softwaredoug | 393 comments on Hacker News.
https://ift.tt/rVe8Qxp...
374 by softwaredoug | 393 comments on Hacker News.
https://ift.tt/rVe8Qxp...
Kimi-K3 Technical Report [pdf]
Kimi-K3 Technical Report [pdf]
380 by vinhnx | 168 comments on Hacker News.
Related: Kimi-K3 on HuggingFace - https://ift.tt/36gMSWZ
380 by vinhnx | 168 comments on Hacker News.
Related: Kimi-K3 on HuggingFace - https://ift.tt/36gMSWZ
Monday, July 27, 2026
Sunday, July 26, 2026
Saturday, July 25, 2026
Friday, July 24, 2026
OpenAI’s accidental attack against Hugging Face is science fiction that happened
OpenAI’s accidental attack against Hugging Face is science fiction that happened
551 by abhisek | 427 comments on Hacker News.
OpenAI and Hugging Face address security incident during model evaluation - https://ift.tt/q96G5Ot - July 2026 (1121 comments)
551 by abhisek | 427 comments on Hacker News.
OpenAI and Hugging Face address security incident during model evaluation - https://ift.tt/q96G5Ot - July 2026 (1121 comments)
Passkeys were invented by engineers with zero understanding of consumer brain
Passkeys were invented by engineers with zero understanding of consumer brain
560 by ksec | 778 comments on Hacker News.
https://ift.tt/HSZFdal
560 by ksec | 778 comments on Hacker News.
https://ift.tt/HSZFdal
Thursday, July 23, 2026
Wednesday, July 22, 2026
Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA
Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA
586 by piotrgrabowski | 326 comments on Hacker News.
Also: Kimi K3: second only to Fable 5 on AA-Briefcase https://ift.tt/1UpQZF6...
586 by piotrgrabowski | 326 comments on Hacker News.
Also: Kimi K3: second only to Fable 5 on AA-Briefcase https://ift.tt/1UpQZF6...
Tuesday, July 21, 2026
Monday, July 20, 2026
Sunday, July 19, 2026
Saturday, July 18, 2026
Friday, July 17, 2026
Thanks HN for 15 years of support and helping me find my life's work
Thanks HN for 15 years of support and helping me find my life's work
412 by nicholasjbs | 39 comments on Hacker News.
Tomorrow is the 15th anniversary of the first day of the Recurse Center ( https://ift.tt/op2qVYm ) My cofounders and I did YC all the way back in the Summer of 2010, with the initial idea of building "OkCupid for jobs." That idea quickly fizzled, and we spent the better part of a year pivoting between other ideas that also failed. Finally, we made something that we wanted ourselves: a self-directed programming retreat, where people built fun projects, contributed to open source, and helped each other become better programmers. After running two small batches, we launched on HN[1] and got an incredible reception. That post on HN helped us reach beyond our personal networks and meet programmers from around the world, many of whom have since become friends. HN brought us the majority of people who came to our next few batches, and in the years since, HN has remained our #2 source of applicants (after word of mouth). Alas, pg's comment[2] on HN when we launched turned out to be prescient: Running free programming retreats isn't a billion-dollar business, but it's still a worthwhile thing to do, and has positively impacted over 3,000 people so far. And 15 years on I still wake up every day excited to keep working on it. So, thanks HN, for helping make the Recurse Center possible, and for helping me find my life's work. [1] https://ift.tt/U7IGcH2 [2] "This sounds like a crazy plan for a startup, I realize, but this is the right sort of crazy. In fact, the way the Hackruiters think about Hacker School is a lot like the way we initially thought about YC: if it doesn't make money, it will at least have been a benevolent thing to do."
412 by nicholasjbs | 39 comments on Hacker News.
Tomorrow is the 15th anniversary of the first day of the Recurse Center ( https://ift.tt/op2qVYm ) My cofounders and I did YC all the way back in the Summer of 2010, with the initial idea of building "OkCupid for jobs." That idea quickly fizzled, and we spent the better part of a year pivoting between other ideas that also failed. Finally, we made something that we wanted ourselves: a self-directed programming retreat, where people built fun projects, contributed to open source, and helped each other become better programmers. After running two small batches, we launched on HN[1] and got an incredible reception. That post on HN helped us reach beyond our personal networks and meet programmers from around the world, many of whom have since become friends. HN brought us the majority of people who came to our next few batches, and in the years since, HN has remained our #2 source of applicants (after word of mouth). Alas, pg's comment[2] on HN when we launched turned out to be prescient: Running free programming retreats isn't a billion-dollar business, but it's still a worthwhile thing to do, and has positively impacted over 3,000 people so far. And 15 years on I still wake up every day excited to keep working on it. So, thanks HN, for helping make the Recurse Center possible, and for helping me find my life's work. [1] https://ift.tt/U7IGcH2 [2] "This sounds like a crazy plan for a startup, I realize, but this is the right sort of crazy. In fact, the way the Hackruiters think about Hacker School is a lot like the way we initially thought about YC: if it doesn't make money, it will at least have been a benevolent thing to do."
AWS: Inaccurate Estimated Billing Data – $1.7 billion
AWS: Inaccurate Estimated Billing Data – $1.7 billion
509 by nprateem | 317 comments on Hacker News.
URL already posted: https://ift.tt/n6W5gi1 I've got an estimated bill for $1.7 BILLION over this month. Normal usage is < $5. Obvs have created an urgent AWS support ticket. Anyone else seeing something like this? Update: Reddit link: https://www.reddit.com/r/aws/comments/1uyuaw7/help_my_bill_s...
509 by nprateem | 317 comments on Hacker News.
URL already posted: https://ift.tt/n6W5gi1 I've got an estimated bill for $1.7 BILLION over this month. Normal usage is < $5. Obvs have created an urgent AWS support ticket. Anyone else seeing something like this? Update: Reddit link: https://www.reddit.com/r/aws/comments/1uyuaw7/help_my_bill_s...
Thursday, July 16, 2026
Wednesday, July 15, 2026
Tuesday, July 14, 2026
Monday, July 13, 2026
Ask HN: Add flag for AI-generated articles
Ask HN: Add flag for AI-generated articles
334 by levkk | 182 comments on Hacker News.
Should HN add the ability to flag articles as AI-generated? This doesn't have to act as a regular flag, i.e., it won't de-rank the article; it could just show up as an indicator, allowing others (like myself) who don't like reading AI-generated text, to skip it. Open questions: 1. Why is the regular voting system not enough? 2. Should HN change in response to the gen AI era? It has been successful not changing fundamentals.
334 by levkk | 182 comments on Hacker News.
Should HN add the ability to flag articles as AI-generated? This doesn't have to act as a regular flag, i.e., it won't de-rank the article; it could just show up as an indicator, allowing others (like myself) who don't like reading AI-generated text, to skip it. Open questions: 1. Why is the regular voting system not enough? 2. Should HN change in response to the gen AI era? It has been successful not changing fundamentals.
Sunday, July 12, 2026
Saturday, July 11, 2026
Muse Spark 1.1
Muse Spark 1.1
404 by ot | 211 comments on Hacker News.
https://ift.tt/YEwDLKP... [pdf] https://ift.tt/oHRfJPz... https://ift.tt/BiUjNq8... , https://ift.tt/QkgwmoZ
404 by ot | 211 comments on Hacker News.
https://ift.tt/YEwDLKP... [pdf] https://ift.tt/oHRfJPz... https://ift.tt/BiUjNq8... , https://ift.tt/QkgwmoZ
GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
450 by scrlk | 363 comments on Hacker News.
https://ift.tt/1HUwZ3n , https://ift.tt/vTY7a4c Prompt: https://ift.tt/TPaFJYO...
450 by scrlk | 363 comments on Hacker News.
https://ift.tt/1HUwZ3n , https://ift.tt/vTY7a4c Prompt: https://ift.tt/TPaFJYO...
Friday, July 10, 2026
New York City to ban deceptive subscription practices
New York City to ban deceptive subscription practices
453 by randycupertino | 228 comments on Hacker News.
https://ift.tt/keSvf59...
453 by randycupertino | 228 comments on Hacker News.
https://ift.tt/keSvf59...
Show HN: Getting GLM 5.2 running on my slow computer
Show HN: Getting GLM 5.2 running on my slow computer
724 by vforno | 176 comments on Hacker News.
A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten from models like Claude or GPT, and this really surprised me. But then I thought, "I wonder how it would work on a normal computer like mine," and above all, "I wonder if it would work without going into OOM on a computer like mine." So I started working with the help of agents to test this possibility. I started converting the model to int4, understanding MTP usage, and if possible implementing DSA for long context. How it responds in int4 and whether the quality is maintained or not. Until I got to the point, on my computer with 32GB of RAM, I was able to communicate with GLM 5.2 with times that, of course, aren't high in cold start, but even then, we're talking about 0.1 tok/s, but that wasn't important to me. The important thing was the journey to reach this goal. I just wanted it to work at all costs, even slowly. So I created Colibrì, which was born from a very simple idea, to be honest, but tested in every way, where a 744B Mixture-of-Experts model activates only ~40B parameters per token—and only ~11 GB of those change from token to token (the routed experts). So: The dense part (attention, shared experts, embeddings—~17B params) stays resident in RAM at int4 (~9.9 GB); The 21,504 routed experts (75 MoE layers × 256 experts + the MTP head, ~19 MB each at int4) live on disk (~370 GB) and are streamed on demand, with a per-layer LRU cache, an optional pinned hot-store, and the OS page cache as a free L2. The engine is a single C file (c/glm.c, ~1,300 lines) plus small headers. No BLAS, no Python at runtime, no GPU.No GPU or serious hardware because I don't have that hardware so I can't test it on hardware that is more powerful than my computer.Colibrì is a one-person project, written and tested entirely on a 12-core laptop with 25 GB of RAM — the numbers above are the ceiling of what I can measure at home. Any feedback is welcome! (and if anyone wanted to participate in the project I would be delighted) Repo: https://ift.tt/hRPLpBF
724 by vforno | 176 comments on Hacker News.
A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten from models like Claude or GPT, and this really surprised me. But then I thought, "I wonder how it would work on a normal computer like mine," and above all, "I wonder if it would work without going into OOM on a computer like mine." So I started working with the help of agents to test this possibility. I started converting the model to int4, understanding MTP usage, and if possible implementing DSA for long context. How it responds in int4 and whether the quality is maintained or not. Until I got to the point, on my computer with 32GB of RAM, I was able to communicate with GLM 5.2 with times that, of course, aren't high in cold start, but even then, we're talking about 0.1 tok/s, but that wasn't important to me. The important thing was the journey to reach this goal. I just wanted it to work at all costs, even slowly. So I created Colibrì, which was born from a very simple idea, to be honest, but tested in every way, where a 744B Mixture-of-Experts model activates only ~40B parameters per token—and only ~11 GB of those change from token to token (the routed experts). So: The dense part (attention, shared experts, embeddings—~17B params) stays resident in RAM at int4 (~9.9 GB); The 21,504 routed experts (75 MoE layers × 256 experts + the MTP head, ~19 MB each at int4) live on disk (~370 GB) and are streamed on demand, with a per-layer LRU cache, an optional pinned hot-store, and the OS page cache as a free L2. The engine is a single C file (c/glm.c, ~1,300 lines) plus small headers. No BLAS, no Python at runtime, no GPU.No GPU or serious hardware because I don't have that hardware so I can't test it on hardware that is more powerful than my computer.Colibrì is a one-person project, written and tested entirely on a 12-core laptop with 25 GB of RAM — the numbers above are the ceiling of what I can measure at home. Any feedback is welcome! (and if anyone wanted to participate in the project I would be delighted) Repo: https://ift.tt/hRPLpBF
Thursday, July 9, 2026
Wednesday, July 8, 2026
Tuesday, July 7, 2026
Monday, July 6, 2026
GPT-5.6 Sol Ultra will be in Codex
GPT-5.6 Sol Ultra will be in Codex
383 by mfiguiere | 339 comments on Hacker News.
https://ift.tt/UfAsl5X , https://ift.tt/cXkhHRP
383 by mfiguiere | 339 comments on Hacker News.
https://ift.tt/UfAsl5X , https://ift.tt/cXkhHRP
Sunday, July 5, 2026
Saturday, July 4, 2026
An American Privacy Emergency
An American Privacy Emergency
404 by flowercalled | 134 comments on Hacker News.
Recent and related: Noise infusion banned from statistical products published by Census Bureau - https://ift.tt/up7aikA - June 2026 (604 comments)
404 by flowercalled | 134 comments on Hacker News.
Recent and related: Noise infusion banned from statistical products published by Census Bureau - https://ift.tt/up7aikA - June 2026 (604 comments)
Friday, July 3, 2026
Thursday, July 2, 2026
Wednesday, July 1, 2026
Subscribe to:
Posts (Atom)