Show HN: Wallfacer – A terminal session manager for Claude Code, and more
3 by pradiptasarma | 1 comments on Hacker News.
Hack Nux
Watch the number of websites being hacked today, one by one on a page, increasing in real time.
New Show Hacker News story: Show HN: CityEdit – Open-source map for voting on NYC street changes
Show HN: CityEdit – Open-source map for voting on NYC street changes
2 by edbltn | 0 comments on Hacker News.
CityEdit (cityedit.org, https://ift.tt/QvsNOig ) is a map where anyone can propose street changes, e.g. safer crossings, new bike lanes, improved tree beds. Neighbors can then vote on those proposals. We just soft launched the app by hanging posters with QR codes at NYC's statistically most dangerous intersections (from open crash data at https://ift.tt/BGLSUKs... ). Read more at https://ift.tt/DW7QlIn... (A poster turns out to be a great distribution channel - it reaches exactly the people who use that street, when they are using it!) The ultimate lofty vision is to build a sensorimotor loop that can awaken a city's collective consciousness: the community perceives its desires, proposes, votes, acts. Two disclaimers: 1. A vote on a map is more of a signal than a mandate, and may be unlikely to inspire government action. Instead, for now, we are going to enact what changes we can through some cheeky tactical urbanism. ( https://ift.tt/7Rqlhjm ) 2. There is selection and presentation bias inherent to how we've built this. Our aim is to be as transparent as possible about who is participating, and who is missing, rather than faking general consensus. We're also seeding maps with existing commute data to fill in gaps wherever we can (e.g. Citibike trip data on the NYC Bikes map) Critiques are welcome, especially skeptical takes on whether this type of civic feedback loop can work. But please be constructive too!
2 by edbltn | 0 comments on Hacker News.
CityEdit (cityedit.org, https://ift.tt/QvsNOig ) is a map where anyone can propose street changes, e.g. safer crossings, new bike lanes, improved tree beds. Neighbors can then vote on those proposals. We just soft launched the app by hanging posters with QR codes at NYC's statistically most dangerous intersections (from open crash data at https://ift.tt/BGLSUKs... ). Read more at https://ift.tt/DW7QlIn... (A poster turns out to be a great distribution channel - it reaches exactly the people who use that street, when they are using it!) The ultimate lofty vision is to build a sensorimotor loop that can awaken a city's collective consciousness: the community perceives its desires, proposes, votes, acts. Two disclaimers: 1. A vote on a map is more of a signal than a mandate, and may be unlikely to inspire government action. Instead, for now, we are going to enact what changes we can through some cheeky tactical urbanism. ( https://ift.tt/7Rqlhjm ) 2. There is selection and presentation bias inherent to how we've built this. Our aim is to be as transparent as possible about who is participating, and who is missing, rather than faking general consensus. We're also seeding maps with existing commute data to fill in gaps wherever we can (e.g. Citibike trip data on the NYC Bikes map) Critiques are welcome, especially skeptical takes on whether this type of civic feedback loop can work. But please be constructive too!
New Show Hacker News story: Show HN: Humor Arena – Which frontier model is funniest?
Show HN: Humor Arena – Which frontier model is funniest?
4 by killiandunne1 | 4 comments on Hacker News.
What if you could measure humor? Well we've trained a model on our own dataset of ~50k human ratings to detect what jokes people find funniest. We know it's part objective, part subjective component. Subjective is out of our depth for now haha The main results: Fable 5 is funniest - beating the average model 67% of the time, with GPT 4o last at 17%. Other findings: - The models never refused to try, even with dark prompts - Thinking longer has a slight benefit - Absurdness correlates negatively with joke quality Some methodology notes: - We benchmarked our model against the human majority and it agreed 72% of the time in a blind sample test. - We had 51 US adults rate the jokes, each blind to the models, with joke order randomized, and quality checked for attention and speed. - To rate some yourself visit https://pair.laugh.so The full benchmark here: https://ift.tt/qfS9nEQ Am taking requests if there's more research you want to see! Cheers
4 by killiandunne1 | 4 comments on Hacker News.
What if you could measure humor? Well we've trained a model on our own dataset of ~50k human ratings to detect what jokes people find funniest. We know it's part objective, part subjective component. Subjective is out of our depth for now haha The main results: Fable 5 is funniest - beating the average model 67% of the time, with GPT 4o last at 17%. Other findings: - The models never refused to try, even with dark prompts - Thinking longer has a slight benefit - Absurdness correlates negatively with joke quality Some methodology notes: - We benchmarked our model against the human majority and it agreed 72% of the time in a blind sample test. - We had 51 US adults rate the jokes, each blind to the models, with joke order randomized, and quality checked for attention and speed. - To rate some yourself visit https://pair.laugh.so The full benchmark here: https://ift.tt/qfS9nEQ Am taking requests if there's more research you want to see! Cheers
New Show Hacker News story: Show HN: Modern Browsers Don't Need the Cookie Anymore
Show HN: Modern Browsers Don't Need the Cookie Anymore
11 by kuberwastaken | 9 comments on Hacker News.
11 by kuberwastaken | 9 comments on Hacker News.
New ask Hacker News story: Ask HN: Dear Anthropic, can we please have thought traces back?
Ask HN: Dear Anthropic, can we please have thought traces back?
2 by exabrial | 2 comments on Hacker News.
Dear Anthropic, can we please have thought traces back? I can't verify whether or not the LLM is arriving at the conclusion from cheating, or if it's fudging or making stuff up. Opus 4.6 remains the best model because of this.
2 by exabrial | 2 comments on Hacker News.
Dear Anthropic, can we please have thought traces back? I can't verify whether or not the LLM is arriving at the conclusion from cheating, or if it's fudging or making stuff up. Opus 4.6 remains the best model because of this.
New Show Hacker News story: Show HN: My tool scanned 256 AI-built apps and most had exposed credentials
Show HN: My tool scanned 256 AI-built apps and most had exposed credentials
3 by kahlilashanti | 0 comments on Hacker News.
All of a sudden everybody wants to build with AI. People in hell want ice water. Built something with AI and don't know if it's secure? Try necktochoke.
3 by kahlilashanti | 0 comments on Hacker News.
All of a sudden everybody wants to build with AI. People in hell want ice water. Built something with AI and don't know if it's secure? Try necktochoke.
New ask Hacker News story: Ask HN: What if there was an 8th day of the week?
Ask HN: What if there was an 8th day of the week?
4 by cyanregiment | 1 comments on Hacker News.
Would it be another work day, weekend, or something else? What would it be called? Where would it fit nicely in the week? Sun Moon Tyr Wodan Thor Freya Saturn How would this change things in the modern world? Is there actually an optimum number of days in the week or is it just random?
4 by cyanregiment | 1 comments on Hacker News.
Would it be another work day, weekend, or something else? What would it be called? Where would it fit nicely in the week? Sun Moon Tyr Wodan Thor Freya Saturn How would this change things in the modern world? Is there actually an optimum number of days in the week or is it just random?