Hacker Newsnew | past | comments | ask | show | jobs | submit | chenzhekl's commentslogin

Unless Japan raises interest rates or tightens fiscal spending, no amount of intervention will work as long as the structural problems remain.


But it's priced the same as frontier models. Why do I not directly pay for frontier models?


This is a charitable read, but I think that being able to pick from a panoply of models will actually yield much better results in the long run.

The same model that has been post-trained to operate for hours as a Linux admin will be incapable of writing a heartfelt email, but with something like Fugu, you'd get both the Linux admin for driving the browser harness and the smaller writing specialist model for drafting the email itself.


u get a pool of them + sakana?


I probably will never pay to Sakana, as they are involved in military contracts.

https://japannews.yomiuri.co.jp/politics/defense-security/20...


Yeah, I was trying to parse their "defense policy" https://sakana.ai/company-info/defense-policy.html?lang=en But it seems like lot of words to say we have no policy and we'll just go along with the powers that be. Like they rely on deferring to the Pacifist constitution, which the current administration if moving mountains to try and change. And when it it you can bet they will not want to give up their defense contracts.


Anthropic was very much willing & involved in U.S. military contracts before the falling out with the DoD. OpenAI is actively involved.

https://openai.com/index/our-agreement-with-the-department-o...


I imagine if it was Deepseek partnering with the CCP it would be different?


I was just stating facts about Sakana, and that was enough to trigger you? For the same reason, I don’t use GPT either. At least for now, DeepSeek has no ties to the defense sector. And don’t talk as if the CCP were the devil. The U.S. president is the world’s biggest arms dealer, after all.


Of course DeepSeek is used by the PLA, why would you think otherwise? https://www.scmp.com/news/china/military/article/3303512/chi... There can be multiple devils, after all.


> At least for now, DeepSeek has no ties to the defense sector.

Like every company based in China they are under the control of the Chinese state, which is an armed entity known to use violence.


Why is this on the front page of Hacker News? Isn't it just more vibe-coded garbage where nobody takes responsibility for the resulting code?


maybe you are not bottlnecked by coding. but there is high probability that you will be bottlenecked by verifying the correctness of LLM-generated code.


Crazy how this doesn't register in people's heads. Has the real bottleneck ever been code written and not the review of code and everything involved? Understanding the nuance and implications behind design decisions; strategy.

In any REAL, workload, with good processes, code review makes speed of code generated a moot point. You still move as fast as you can review the code, and no, I won't debate that you can rely on LLMs, a deterministic language predictor, to determine the correctness of code; in the context of the business, and technical implications.


That is indeed the point I was making.


Where is the real bottleneck, if I may ask?


> verifying the correctness of LLM-generated code

It's... pretty clear in the original conversation.


I find that people who write "may I ask" are often/usually bad-faith arguers under cover of being polite.


That's a good rule of thumb, it seems that way more often than not.


If you are a responsible maintainer you need to verify the correctness of the contribution wether you used an LLM to generate it or wether someone else did.

Having someone else be the AI-middlemen, just introduces additional complexity and confusion.


It's interesting that they mentioned in the release notes:

"Limited by the capacity of high-end computational resources, the current throughput of the Pro model remains constrained. We expect its pricing to decrease significantly once the Ascend 950 has been deployed into production."

https://api-docs.deepseek.com/zh-cn/news/news260424#api-%E8%...


Yup, I tried to benchmark it, but harder questions time out or get rate-limited...


Sorry, but exactly where in the article that you linked contains the mention of " Ascend 950"?


it's in the footnote text of the first figure of the section the link points to, where "昇腾950" means "Ascend 950"


OK, strange that it doesn't appear on my version of the webpage

https://api-docs.deepseek.com/zh-cn/news/news260424#api-%E8%...

This is the first figure of the section that the above links point to (https://api-docs.deepseek.com/zh-cn/img/v4-spec.png).

And I can read Chinese.



It feels like the current trend is a bit scary: the more AI advances, the more people with money and resources will gain disproportionately greater advantages. For example, they can make their own software more secure, while also finding it easier to discover ways to attack other software.


You can already do that today by hiring a security researcher. I can guarantee you that Apple has access to people of a higher caliber than my startup.

I could see a world where 1 year from now I can have glassing do a full sweep of my codebase for a given price (say: $10k). Running that once a year is within my means and would make my software much more secure than it is today.


Yeah but even Carlini who is a good security researcher said he has found more valid vulnerabilities in the last week than his entire career before this. That sounds like it’s clearly better/faster/cheaper than a human security researcher that would cost $300,000 a year.


I spend well over that of my employers money on pentesting every year. I’m absolutely certain Claude could perform as good or better a job using what’s available today.

It had crossed my mind that an AI agent pentester would be an interesting product to build. Once again though, the labs are just going to build it because it’s a thin thin wrapper.

Beyond existing software with vulnerabilities, the really important aspect of this for Anthropic et al is that the gigatons of code that are being generated every day needs to be secured.


There are quite a few such startups already out there. Results are mixed so far. Though I believe they get much better over the coming months and years.


AWS has one as a managed service.


Sounds normal to me!

i.e. it may be a step change and that could very well have distinct and noticeable real world effects, like other technologies have in the past, but it’s nothing fundamentally new.


It only feels like that if you’re just catching up. The logical consequences you are just realizing are the reason OpenAI was founded.


This has increasingly been my take. If we accept that AI is an amplifier of impact, then it follows it will amplify disparities.


yes, this is what am afraid of, the gap is going to increase more as AI advances further.


My impression is that, before Microsoft acquired GitHub, GitHub went for many years without really introducing new features, so part of its stability came from the fact that it wasn’t very ambitious or proactive about improving.


I loved that time. Websites, or "apps" that don't change every second time I want to use them, are great.


The statement from OpenAI makes me feel that Sutskever was right; Altman is full of lies and will say anything for his own interests.


This tells us that we should never share sensitive information with GPT, even if you’ve set it not to use your data for training. Nothing can stop OpenAI from misusing your data.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: