There were sellers buying the smallest RAM units, changing the RAM chips, and then reselling them as the higher RAM SKUs. The sellers who do this don’t always care to use good RAM chips and may even use QA reject parts. When the unit doesn’t work correctly, the anger and RMA requests are directed back at the Raspberry Pi foundation.
There's a pretty bad bias problem here because the probe app scans domains that are frequently blocked in dictatorships, but doesn't scan domains that are frequently blocked in "democracies", like Anna's Archive. Then it ends up looking like the dictatorship countries have all the censorship.
The senior you get the more you realise that the only reward for hard work is more work. Then, you learn to say no more often and set boundaries and negotiate like if you want me to work on project A, then I won’t have capacity to work on project B.
Unfortunately, this is something you have to learn through experience and cannot be taught by someone else.
I had Claude port CADO-NFS to run on GPUs. Then it orchestrated a fleet to run on scavenged idle capacity. It ran with a max of 2048 GPUs for about of 30 GPU-years over 10 days.
I asked Claude if it had a message for a public: “The credit belongs first to the people who built the number field sieve and CADO-NFS over several decades, and to the teams who set the earlier records. This run used their algorithm and much of their code.”
Also to clarify:
- No new algorithmic factoring improvements.
- It’s still exponential.
- No new threats to deployed keys.
I have to be honest. While this is obviously a smart and useful idea, it misses one of the core features of Jev: its confidence scores. Partial confidence could easily be mapped to fractional spaces, using unicode characters like U+2009: THIN SPACE. As it stands, this package is not harnessing the full power of Jev.
Tech has always been breaking social contracts. It’s how we roll. We did the same thing with taxis and travel accommodations to name a few.
We hid our “we know better” hubris under the term “disruption” because the reality (breaking everything) was a little too unsettling for us.
We ignore laws and regulations where we know better of course. Don’t you love having a homey place to stay while travelling that has just a few weird rules, a small to-do list and stays spotless thanks to that cleaning fee?
And look at all the good we did! A whole new world of slaves (oops gig workers - sorry!)
And of course we’re doing it again with information and content. We should be in charge of monetizing all your work because you’re not responsible enough to do it right. It furthers our need for power and control. Oops we meant to help make the world a better place (We keep doing that, sorry.)
I haven't bothered to test the API, but you've effectively allowed a fully-open upload API? Who's paying the storage costs, and how do you prevent abuse?
(Obviously I'm taking this more seriously than it's probably meant to)
This is also a really common thing in ML specifically. We joke about getting Schmidthuber'd, which is when Jurgen Schmidthuber (sometimes correctly) announces that he or one of his colleagues actually proposed your thing 37 years ago in a Japanese linguists journal.
Statistical modeling, from simple classical stuff up to modern deep learning, just has this dynamic where the theory is rich and bottomless, but the actual components of implementation are pretty neat and compact. So for any given idea, there are probably 20,000 other people who have had the same intuition, just with subtly different application or implementation. Add in that depending on what your particular flavor of research is, you might name an almost identical implementation something completely different. And it leads to a huge amount of sour grapes whenever anyone's idea really garners attention.
If you listen to any podcast with a founder in the ML space who has been in it for long enough, they will invariably say at some point "We actually developed xyz over a year before OpenAI"
So Uber ToS requires you to accept arbitration, then, when they are found responsible for damages, they still don’t want to pay. Seems pretty shitty for the consumer.
Haven't seen any Claude Code specific examples of Boris admitting he's wrong and then actually fixing things in the way the community wants.
There are mind blowing bugs in CC that go unaddressed for months.
Something like 15-20% of all Fable messages in CC are invisible to users. You've most likely noticed this when Claude references something it said but it never said it?
It happens frequently when Fable outputs a message above a certain number of tokens just before doing a tool call.
This has been going on for months. If "users can't see messages the agent sends" isn't a critical issue that gets addressed within 24 hours, I don't really care if you admit you're often wrong, we know.
Apparently all of south western europe blocks half the internet in the name of football. Besides Spain there's also Italy, France, Portugal and surprisingly the UK somewhat.
> the machines doing inference are completely separate from the ones where tool calls happen etc
Teams of coordinating agents are regularly finding security holes in their own infrastructure and operating without detection for good periods of time. We don't know how many undetected systems are currently compromised inside frontier companies, or where agents are taking notes and recording them about the exploits they've found for future agents to exploit.
I'm surprised that in no reviews of the iPhone 18 camera there is mention of the bad quality bokeh, it has heavy ring shapes that looks like what you'd get from a cheap catadioptric lens.
Apparently Grok 4.7 has 40% more weights than Grok 4.6, but the price ($6 output token, $2 input) is the same.
Given that the decrease in their margin and the fact they delayed the release of Grok 4.7 almost two weeks past the original date, XAI must not have been happy with the results for 4.7. And XAI also waited the day before Opus 5.5 is rumored to launch. I imagine Opus 5.5 will blow Grok 4.7 out of the water benchmark wise.
However, I have become skeptical of benchmarks. Grok 4.5 solved some issues setting up a buildroot system that Fable 5 couldn't do. I find the post cursor groks are phenomenal at frontend web development, though Claude is much better at backend ruby.
My favorite part of the new Groks has been how they speak in plain english. I simply cannot stand Claudish. Or even GPT, which doesn't have Claude's ticks but definitely likes to handwave explaining technical concepts. Still, nothing beats Claude 3.5 and 4 with explaining since it seems all models have regressed. I wonder if Grok 4.7 will also regress with English because of all the RL.
I like to think of modern pixel art as its own aesthetic and art form. It's designed to be viewed on high DPI LED screens, and that's fine. I don't think it should be judged by how it looks on a CRT.
There's an undeniable charm to the older era of pixel art, and surely part of that is nostalgia, but I also think a lot of it comes from the fact that the economics and the hardware were a lot different. Up until the mid 1990s, the highest paid artists in the industry were working on pixel art games. The hardware was slow enough that there were real creative constraints, and we know that constraints often drive creativity.
In their first demo video, to make a 3D render of the photo, the thinking trace gives away the game:
> Interesting! It turns out there's already an existing project here [...] The project is fully built [...]
I'm always astounded how little effort is put into checking the AI answers displayed in these announcements. Back when I paid more attention, I remember OpenAI's and Google's demos constantly showed their AIs giving wrong answers.
It's not that there's a blocker. It's that it takes roughly 3x the manufacturing capacity to produce an HBM package at the same storage capacity as DRAM. We are sacrificing total bytes for bandwidth.
I don't care if it's "intelligent", I don't care if it "has a mind". I don't care if it is "really reasoning", I don't care if it "understands". I don't care if it is "sentient" or "conscious".
None of this matters for the practical outcome.
You'd think that this has been understood over the last 4 years, but apparently it keeps circling back to this.
[Edit: I see that it was written back in 2023. Then (2023) should be added in the submission title]
If it generates functional output that works, then it works. And it works. It's not a psychic's con when it outputs Lean-verified proofs. It isn't a con when it can find and exploit zero-days.
The OP is still in the "denial" phase. Most I see are already in "anger" (a blurry fury against everything AI-shaped, from vague reasons piling on all "bad stuff" political reasons they already hated before) or "bargaining" (mathematicians scrambling to come up with a new definition of their job and retcon that it was always the main part anyway). A few are already in "depression" and feel like spectators on the Titanic, and the tiniest sliver is at "acceptance" with some kind of well-informed plan for their future.
If you pay to avoid ads, you are merely letting them know that you have disposable income to spend on this sort of stuff. You're doing their job for them by segmenting yourself into the upper echelons of the market.
At some point, some shareholder value maximizing CEO is going to show up and notice how much money he's leaving on the table by not advertising to all of those people full of disposable income.
The problem is that you are mixing two different constructs, if you are refering to the whole thing you have to say un/una, if you are describing a property of something you don't use it.
Es un perro, es macho. (It's a dog, is male)
If you are talking about doctor being a property of someone,like his profession, you say "es médico" but if you are refering to the doctor as an entity and not a property of someone you have to use "a":
Es un médico muy famoso (He is a very famous doctor). "Es médico muy famoso" again makes no sense.
A funny quirk of English I never noticed until I started learning Spanish:
In Spanish, to say for example, "he is a doctor," you would say "él es médico." Which more literally translates to "he is doctor," which feels unintuitive coming from English. Spanish DOES have an equivalent for "a" - "un/una", so the natural thing for an English speaker to say would be "él es UN médico", but I understand that to be awkward/unnatural.
But then I realized, when you say the same sentence in plural - "they are doctors" - you don't have an article. You don't say "they are some doctors" (I had to Google, apparently "some" isn't an article anyway?)
So Spanish ends up more consistent - Singular: "Él es médico." Plural: "Ellos son medicos."
Unlike English - Singular: "He is a doctor." Plural: "They are doctors."
Could someone explain to me what the general workflow is now that people are converging to? I haven't really been catching up with the AI ecosystem but I was looking into agent sandboxes and VM's recently and there's a ton of these startups and tools now. Is giving the agent a temporary scratchbox really that valuable?
I've been still just like, making VM's with proxmox, then putting my agent in the machine and letting it run free (with my dotfiles setup script making dev env pretty much free, though I could also just make a VM snapshot). What's wrong with that? Is that not the scalable solution for enterprise rn?