Hacker Newsnew | past | comments | ask | show | jobs | submit | codexon's commentslogin

How did you scrape reddit comments? Doesn't this require expensive licensing from reddit? How do you handle comments that were deleted by users?

pushshift baby

How hasn't reddit sued them offline yet?

The problem is most of the botnet zombies are in places with very little rule of law like eastern europe/russia/south america/china. It is an exercise in futility and the richest customers just opt to just spend money on more protection than shutting down the zombies.

If the zombie is in the US, hosts like Google or Amazon will take 1+ month to respond.


Then IP-range block those places until they get their affairs in order.


Author here. This attack was extremely broad. I saw parts of this attack come from my own home ISP's ASN, though not my IP thankfully. If we just "blocked those places" there would be a lot of collateral damage. As it stood, we did temporarily bump up rate limiting for the biggest attack ASNs and we absolutely heard from real, regular users about it (Sorry to those affected).


Either it's coming from certain places and you can block those places, or it's coming from everywhere including some that are within your legal jurisdiction.

The wafer only has space for 44 gb of sram. If they offload ram they lose the speedup of having everything on 1 chip (the whole point of cerebras).


They can host larger models by pipelining it on multiple wafers. Each wafer stores one layer and N layers can serve an N * 44 gb model with N concurrency. The limitation would of course be inter-wafer I/O, which my comment was getting at. That's probably how they can serve bigger models like GPT 5.6 Sol [1].

[1] https://www.cerebras.ai/blog/accelerating-gpt-5-6-sol-ultraf...


I never said offloading was impossible. It will result in a large slowdown.

It would look bad for cerebras if other people are hosting the 27b version and show a higher TPS than cerebras.


They are still offering full mythos to project glasswing companies and those that pay them enough.


It isn't just about money, it's about who. I don't think the companies using it are using it to create botnets.

The intention is for highly targeted pieces of software to use it to secure their code and be ahead of the game before the open market gets access to the same capabilities for offense.


They gave a few of the largest companies access to mythos.

Half a year later, it is still not available to everyone else.


I don't have any issues with the extra features they try to tack on. I just disable them.

I just wish they would fix the memory leaks and stuttering.


I want to be able to recommend firefox to everybody, but after 15 years of chrome and going back to firefox, it still feels jank.

Web apps like twitter youtube discord that is left open for a few hours take up >2gb ram and is jittery. I have to shift+esc and close them to make them run smoothly again. Why hasn't stuff like this been fixed yet?

Why do I need to open about:profiles just to launch a new profile?


I heard the only way you are going to get into the CVP program is if you have public CVEs. Doesn't seem to matter if you are in a company account or not according to people that are supposedly in the program.


Thank you, that's actually helpful. I thought it was only accessible through a b2b agreement with Anthropic.


Rumors are that if your company spends a lot, they will also remove safeguards.


I used to use firefox a long time ago but got tired of all the memory leaks and lag before I switched to chrome.

Now that mv2 is going away, I switch back to firefox, and it still has memory leak issues after all these years, it is really disappointing.


I'm not a biochemist and I have been blocked by fable and opus for "cyber" just for doing things like asking it to ssh into one of my servers, look at a CVE, or do work in assembly. Been rejected to their cyber verification program 3 times already.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: