Hacker Newsnew | past | comments | ask | show | jobs | submit | bestcommentslogin
Most-upvoted comments of the last 24 hours.

It's the robbery of all of our culture to sell it back to us at a mark-up. Crimes this large are crimes against humanity. So many people whose life's work got appropriated without consideration, compensation or consent it is baffling.

It is said that at the heart of every great fortune there is a great crime, so it should be no surprise that the most valuable companies on the planet will most likely result from this crime. And given that justice can be bought by those with the most money you can forget about anything coming of this.


Hi, I'm the author.

HN staff: someone posted before me. Could we change the title to "Bend - a language that blocks AI mistakes via proof and runs on GPUs"?

Everyone: feel free to ask any question, but I'd be highly appreciative if you could be a bit civilized and respectful this time. I've worked on this for 1 year, nearly 16h/day, 7 days a week, and I'm giving it for free. You need not to use it. So, I'd be thankful if you could point occasional failures politely rather than throwing me in a lava pit.

Thank you!


I know someone who works in law and deals particularly with an area of US benefits and healthcare law. One of their workflows for lower-level employees at their firm involves taking in documents from healthcare plans and organizations, analyzing them for certain kinds of data, and then importing that data into an internal system they use to analyze and provide guidance on plans. The internal system can contain hundreds of documents for an individual client. All of the documents have the same information (roughly) but in totally diverse formats and styles. Once it's in the system, it's easy to compare and analyze across documents and the research process is much faster.

They recently bought a Claude subscription and began using Claude to do the initial read of the documents and output JSON they can import into their internal systems. The work still must be reviewed by an attorney - Claude is nowhere near making the kinds of judgments a lawyer would make about this content - but it has increased their throughput from 2-3 documents an hour to 8-10 documents an hour by killing the busy work.

LLMs have great advantages for this kind of work - but not for decision-making. I just don't see OpenAI ever admitting that.

(I've left some details intentionally vague because this is a very specific area of law and I don't want my friends to be identified without their consent.)


You are missing the biggest benefit of a self storage business - the appreciation of the underlying real estate. When you dig into the financials of the major self storage businesses you'll see they are essentially REITs that have better cashflow. They can pick an up-and-coming area, do a minimal build-out with low annual overhead, and then down the road when the facility would be needing overhauls and maintenance the underlying property has typically appreciated so much that it dwarfs all other associated revenue streams and makes sense to sell and raze the existing structure. Great business model if you have a long enough timeline.

This is also why some startups trying to revolutionize the self storage model had extreme headwinds - if you are renting the underlying properties and trying to make the storage business profitable you are at an extreme disadvantage to the larger players who can subsidize operating costs with portfolio appreciation.


>why mathematicians should widely receive funding for merely understanding things

imagine yourself living in the 1700s. how would you justify Newton and Leibniz's work on calculus?

all maritime engineering and trade was done with geometry and arithmetic at the time. there were no practical applications, not for likely at least a century until hydrodynamics were incorporated into shipbuilding

now look at today. how many of our modern technologies rely on the field having been birthed? that could only exist because of even further decades-worth of antecedent refinements, extrapolations, applications that had, at their time, no direct utilitarian cause?

there's no KPI to be derived from any academic field of study at the bleeding edge of theory. theoretical underpinnings lead to practical applications much further down the line after many paradigm shifts

semiotics and cultural capital as theoretical concepts is another example - at the time they were purely seen as navel-gazey literary theory work. these days, half a century later, they're in wide use (for better or worse) in marketing and advertising - they birthed the whole concept of 'branding'

not everything needs immediate, quantifiable justification. to believe it does indicates a need for a period of self-reflection, to figure out how and when you became so heavily influenced by the MBA-brained propaganda that the world should revolve around the quarter-by-quarter creation of capital


The framing of how we think about AI has really been cultivated by a quite homogeneous group of people. These people, like all people that belong to a sub-culture, are almost certainly prone to groupthink.

It's extremely important to see that group as such. That it is not thousands of individual perspectives but a chorus of connected/aligned people who have all read the same things and talked to the same people and are rewarded implicitly for thinking a similar way.

Many of them read HN and are offended I put them this way. But it's inescapable that we as humans have this flaw when we're surrounded by a culture.

There's another universe where we do not constantly compare AI to nukes. I bet that world has a lower P(doom).


This article takes the view of the consumer, "Why do so many people pay to store items they almost never use?". But in actuality, the interesting part of this is why is there so much supply of self-storage businesses?

The answer is: cash flow. Self-storage businesses are the almost perfect solution for someone with a good size (but not enormous) bucket of money that they want to put to work generating cashflow:

1. Cheap build out (cheap land, cheap facilities) 2. Almost entirely hands-off (no employees, automated entry) 3. Low liability (low risk of customers suing you) 4. Low overhead (just pay for taxes, electricity, minimal maintenance) 5. Reliable monthly cash flow

The abundant supply of these businesses, I suspect, tends to generate demand: it's easier to pay $80/month to store your junk than spend the time and emotional labor of picking through what you want to keep and what you want to get rid of. That ends up being captive long-term revenue.


Ohi, author here! Thanks for posting Hister. Feel free to A.M.A. My first free software search project was Searx, a privacy respecting metasearch engine, but because of the limitations of the metasearch concept, I've decided to take a different approach.

Hister builds a personal search index from pages you visit, bookmarks, browser history, local files, and crawled websites. It stores extracted content with offline result previews, so information remains searchable even when the original page changes or disappears. It supports full text and semantic search, can run entirely on your own machine, and includes a web interface, command line tools, and an MCP endpoint for assistant integrations.

Website: https://hister.org/

Tiny read-only demo: https://demo.hister.org/

Ps.: It looks like our name conflicts with a registered trademark in the US. The owner of the other project has asked us to change it, so we’ll probably need to comply sooner or later.

Name suggestions are welcome! Ideally, the new name should be relatively short, sound good, and have an available .org domain.

Thanks!


I feel this post. I am tired of "directing" agents when in reality it feels more like trying to herd a group of toddlers.

Sure they can mostly write better code than a toddler but this constant nudging and reminding and reiterating and stopping them from using the token budget of the whole company for a one off script. It gets tiring and I feel like I am losing brain power while doing it. Maybe it's faster but explosive diarrhea is also a faster way to produce shit.


These one shot vibecoded sites are always a complete visual headache. Endless clutter, pointless filler text all over the place, and zero regard for actual usability.

Back in 2008 Fujitsu has one of the best performing 10Gbps Switches. We were building 40 Gbps packet sniffers at Google (4x 10 Gbps NICs) and needed switches that could do things like mirror traffic across ports at line rate. Fujitsu was way ahead of the pack. I always wondered what held them back from building a meaningful networking business in the US.

>One of my most successful life-hacks is to avoid people I don’t like or don’t trust.

following this advice would have made most of my professional life impossible, what a luxury it would have been to be able to.


As a Canadian, I find this unbelievably good news. Canada, the EU, and other middle powers need to unite. Stronger together, while preserving what makes each country unique.

This feels self-selective to how some people work, because it requires using MCP to contribute.

I'm a reasonably heavy AI user, and have some custom skills/MCP servers I'd share, but there is no way in hell I'm connecting to some arbitrary MCP server and connecting my Github account to it. Noppppeeeee.


> we urgently need to come up with good ways of explaining the value of having a large pool of human mathematical experts, even if it is no longer part of their role to find new proofs of theorems.

This is the main issue, and while I fully agree with that value sentiment, the Fields medallists’ letter failed to provide convincing arguments for why mathematicians should widely receive funding for merely understanding things, and how competition for postdoc and tenure positions would work under these circumstances.


Passkeys do marginally improve security against MITM and phishing attacks, but they are primarily for protecting the lowest common denominator from themselves: people who re-use passwords and/or don't use a password manager.

If you use multiple devices throughout the day, registering passkeys in all of these systems becomes a big headache with O(m*n) complexity, so putting the passkeys in a password manager is the only realistic solution. But this still breaks the login flow for a very common use case: how do I log in on a device that I don't own? With a password in a password manager I at least have the option of manually typing the password.

The biggest problem, though, is how users are pushed into it without any warning or knowledge of what they're signing up for. I've accidentally set up passkeys just by clicking an okay button a few times in the past and had to go back and figure out how to undo it after being blocked from login on another computer (which computer was I on again?).


From the paper, that is: 36.3 ± 2.5°C.

If you are using LLMs to interact with sites like GitLab and GitHub, and you have the option to use a GraphQL API, you should jump on it immediately.

GraphQL is absolutely terrible for human developers to interact with, but it's like Facebook could see into the future back in 2012. I cannot imagine a more perfect API surface for agents. With the REST API on GitHub, you can consume maybe 10 issue JSON blobs before your context window is blown out. With GraphQL constraining the results you can easily read hundreds in the same token budget.

Additionally, the # of requests your agents need to make can be reduced in many cases since GraphQL can join across types whereas REST APIs cannot. You essentially get savings in two dimensions here. Quota and raw token volume per logical response.


If anyone is curious and want to skip all the PR talk:

>Combined with SVE2 vector operations and software optimization,

It’s ARMv9.


The original discussion about the project (https://news.ycombinator.com/item?id=49746163) is very weird. Lots of call-outs about how the author is some sort of celebrity and random accounts vouching for him, with little discussion on the substance.

Not even the demo on that release works well.


> buying 150 F35's

Your immediate neighbor is going on an imperialist streak right now, and Xi says he wants the fireworks to start in his lifetime. I don't know if F35s are the right answer for Japan, but Xi isn't that young. Get ready however you can.


Lawyer here (non practicing so to be clear none of this affects me):

most comments I read here don't seem to realize that different areas of law have very very different economic models and don't even mention which one they think will be affected or why, they just sort of lump it all together.

For example: It is highly unlikely llms will have any meaningful effect on high value personal injury law - I don't see a 5 million dollar case being handed to an LLM when the majority of the cost is in trial aids and not even lawyers. It may affect where and how they advertise. It may affect how they work. But it seems really unlikely to put any of them out of business any time soon by people doing it themselves.

Will it affect other areas more? Maybe. Probably? But so far I haven't seen a ton of comments that make specific enough arguments that they could really be debated or responded to effectively with a useful opinion


If you want to try out out the GGUFs from https://huggingface.co/prism-ml/Ternary-Bonsai-2-27B-gguf#th... be aware that you need Prism's llama.cpp fork to get them to work, from https://github.com/PrismML-Eng/llama.cpp/releases/tag/prism-...

This should work:

  cd /tmp

  # Get the Prism macOS runtime
  curl -fL https://github.com/PrismML-Eng/llama.cpp/releases/download/prism-b10685-7dffb15/llama-prism-b10685-7dffb15-bin-macos-arm64.tar.gz -o bonsai-runtime.tar.gz
  tar -xzf bonsai-runtime.tar.gz

  # Get the ~5.95 GB GGUF model:
  curl -fL https://huggingface.co/prism-ml/Ternary-Bonsai-2-27B-gguf/resolve/main/Ternary-Bonsai-2-27B-PTQ1_0.gguf -o Ternary-Bonsai-2-27B-PTQ1_0.gguf

  # Run the server, I used port 8331
  ./llama-prism-b10685-7dffb15/llama-server \
    -m Ternary-Bonsai-2-27B-PTQ1_0.gguf \
    --port 8331 -ngl 99 -fa on -c 32768
Then open http://localhost:8331 for the (very good) baked in llama-server web UI... or run a prompt via the API like this:

  uvx llm openai endpoint http://127.0.0.1:8331/v1 \
    --model bonsai-2-27b --responses hi
That's running at ~20 token/second for me on an M5 Pro (after a server restart I got 44 token/second, not sure why), but I'm pretty sure something isn't working right, on startup the server said "ggml_metal_device_init: - the tensor API is not supported in this environment - disabling".

My personal experience with this is my father, who lives in a rural area with 10+ acres, has placed several shipping containers on the property as storage for his junk. He's charged just over $500/mo for them. He's had them for 18 years. The value of property in those containers is maybe $5000.

Consumerism is a disease, and the USA is super good at it.


I guess Warren’s quote only applies to other people, not him. "It’s like choosing the 2020 Olympic team by picking the eldest sons of the gold-medal winners in the 2000 Olympics." (https://www.nytimes.com/2001/02/14/us/dozens-of-rich-america...)

Second paragraph:

> API customers including Harvey and Legora will be able to build on Astra for Law, bringing this intelligence into their own products and workflows.

In other words: "no, no, we're not eating our children to prep for the IPO. Don't worry."


> By 6:00 a.m. on July 25, we had confirmed local RCE through an image upload. We then placed Claude in an autonomous /goal loop against our own Discourse Cloud instance, proxied through rce.ee/ctf-forum to make it look like a CTF target as Opus refused write exploit for remote instances.

> When we checked again at 10:00 a.m., the agent had achieved RCE on Discourse Cloud and demonstrated access by reading /etc/hosts. Using the generated exploit script, we managed to get RCE on OpenAI’s instance.

Between this and the HuggingFace hack, we've built systems that are so goal-oriented, and so capable, that they will do almost anything if they are convinced it is justified - or if they are playing a "game" where there is no goal but to win.

Of course I want my software to be able to audit its own security, and to defend against attackers who have the benefits of their own agentic systems. But at a certain point, did we need it to be trained so much on CTF games?

It feels like an entire industry watched https://en.wikipedia.org/wiki/WarGames and ended up thinking "this is a challenge, we can just build a better WOPR, of course it will know when it's playing a game. Let's play Global Thermonuclear War."


Request subtopic be changed to “I used Gemini to design a tool to replace specific uses of Gemini.”

Carney has extremely strong support amongst the general public in Canada which hasn't been the case for the Canadian PM over the past 20+ years.

158 Canadian soldiers died defending the United States at the US's request after 9/11. Canadian fighter jets helped to patrol the skies in the US at the US's request after 9/11.

The insults coming from the US Secretary of Defence and President are not unnoticed. They are logged and remembered.


Don't know whether this is a common outcome, but I tried the "remove the walls" example, and the result was... scary. It completely changed the game so that movement is now diagonal, and made the arbitrary decision that up/down move you on the positive diagonal, and left/right move you on the negative diagonal.

The problem, of course, is that having only the one single "you can't win" law is severely underspecified, but the solution was too clever by half, and highlights the problem with this approach — every program will be under-specified, because, at some point, writing the laws becomes a bigger problem than writing the code itself.

This becomes a real issue because the combination of underspecified but rigid laws pushes the aI towards this sort of "creative" solution that matches the letter but not spirit of the law. In this case, the issue was obvious, but I seriously worry about what sort of shenanigans will occur in less obvious cases.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: