Hacker Newsnew | past | comments | ask | show | jobs | submit | MASNeo's commentslogin

Calls from law enforcement is laud and clear: BigTech is not doing enough to fight fraud, war and other crimes, they make money from criminals.

Nobody’s yet cared enough to sue them, or their managers at ICC yet. However, eventually an ambitious lawyer will try to make the case, especially if fraud and human trafficking continues to grow like it is.


Hope they don’t rationalize that minimizing paper clips of others is easier and thus do that instead… A more carful person might be scared to even write this on the internet these days, not knowing if it would be the final pin to civilization.

It seems like the author is pitching topological teams and responsible engineer concepts, sprinkled with decision governance.

None of that is really new, but it’s the chaos of silos, narrow decision frames and lack of detailed understanding by decision makers which probably accentuates the adoption of new concepts.


Maybe Musk was right about OpenAI after all?! The ethics are clearly troubling and where something like this pops up there is mich worse that did not made the light of day.

The one time I got to ask Bill a question was probably why Windows didn’t have a “make world” that just works like in FreeBSD. I wonder if he remembered that question during his ordeal.

I wonder if this goes down as AgentGate because clearly HuggingFace was not an isolated incident.

Well worth a material business restriction until an investigation on the root cause by independent parties has concluded and remedial action taken - well, in any other industry but BigTech.


Does anyone feel like everyone chasing the release of Anthropics Fabel 5.1 in a Mad Rush(tm)? In this situation it feels like tuning to benchmarks and other marketing devices feels like trusting Meta in mental health protection of users…

My agents can deal with Excel all right. What’s the story?

The story is that the writer is Orcaset, which is trying to sell you their own analysis tools.

Yup. Their benchmark appears to consist of giving an agent XLSX files that use advanced Excel features, but no copy of Excel, and expecting the agents to interpret the spreadsheet the same way Excel would. So the agents try LibreOffice instead, and if that doesn't work, they try to write Python scripts. But the behavior of the Python doesn't match the behavior of Excel.

To put it politely, this seems like a self-inflicted problem. If you really want your agents to interpret advanced Excel features exactly the way Excel would, have you considered maybe giving your agents Excel?


For as long as criminals believe they can hide in relative anonymity (e-mail, stolen CC and burner phone being sufficient authorization) and high volume of transactions PayPal will remain with us. It’s a great way to do small cash like transactions without going to crypto. Some crimes still need to source cash.


I have had more money stolen from me by PayPal than any scammer.


We mostly use specialized small models because larger are to expensive/slow and are prone to hallucinating quite a bit. Not sure how this is a surprise, seems more like a best practice.


Is there evidence that small models hallucinate less than big ones?


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: