Hacker Newsnew | past | comments | ask | show | jobs | submit | linzhangrun's commentslogin

The speed of light, wave-particle duality, and so on, always make me feel that the universe also has a "computing power"

They say v4.1flash is so strong that they'll route API calls to v4pro to v4.1flash, lol

super fast true


update: Refresh and it won't appear again


Obviously a Fields Medal–level achievement... Even the most optimistic person would not have expected it to happen so quickly six months ago


Under the same computing power, the improvement in LLM intelligence and the improvement in the upper limit of LLM intelligence are equally astonishing. At least in coding, the best local models that can run smoothly on a DGX Spark are now less than one year behind the strongest SOTA models in capability (I measured around 60 tok/s).


Why is the Chinese text missing in many places?


Feels like Gemini Pro will arrive directly as Gemini 4 Pro


Could this be related to the GPU? I remember that sites like Reddit use subtle differences in GPU rendering to fingerprint devices.


I tried it yesterday. The experience seems smoother than on Windows. I hope Computer Use will be added soon. Also, when will OpenAI fix the issue where Xhigh and Ultra are both translated as “极高” in Chinese? It’s been this way for quite some time. As far as I know, the proportion of Chinese employees at these Silicon Valley AI companies is quite high.


You have "send feedback" button in the app. Something tells me that it's much better place to report issues with the app rather than comment section on unrelated website.


I think their backend data probably shows that very few Chinese users are using Codex through the official way, and the number of those who have switched to the Chinese interface is even smaller, so there’s not much incentive to fix this issue.


I don’t think GPT-5.6 would make such a basic mistake when reviewing CodeX itself.


Because all Chinese users must use a VPN to access Codex.


On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half.

Put Codex to work on deploying it now, hoping the speed can improve quite a lot :-) Thanks anyway


> On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half.

That's crazy, a RTX Pro 6000 does that in in 2-3 minutes (give or take, depending on your exact settings). LLMs don't make the difference between standalone GPU vs unified memory + CPU so obvious as diffusion models seems to do.


It’s always been the case, it’s more the anomaly that LLMs work at comparable speeds on M series because almost all other ML runs way faster on Nvidia cards.


LLM prompt processing and diffusion models are compute bound, while LLM token generation is memory bandwidth bound.


An RTX6000 is a completely different class of hardware.


Really? No wonder I keep trying to type on it like a laptop but it doesn't work and doesn't even have a display!


Right? Do you also type on a Mac Studio without a keyboard plugged in? Like tap the ethernet port 3 times in a row then this sequence of sticking your fingers into TB5 ports? I mean, it’s clear you stick your RTX into a computer.


Please keep up posted about the results!


First batch of quick test results: approximately 1/5 speed improvement


+20% or x5 speed improvement?


20%


> Put Codex to work on deploying it now

Which codex?


For gods sakes, Apple let people have run other GPUs instead of these pissweak 2012 class mobile GPUs


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: