No you can't. It still responds in Chinese after explicitly asking it to "Always reason and respond in English."
Was already released before your comment: https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
I use DeepSeek V4 Flash (before this update) for most things: - OpenCode for codebase editing: python scientific computing and LLM projects - Open Interpreter Classic (python version) for Swiss army knife terminal…
DeepSeek-V4-Flash-0731 scores higher than Fable 5 on Terminal-Bench, and Fable 5 is Mythos, correct?
Yes, they trained on my open source code and Wikipedia edits and Stack Overflow answers that I shared freely, so it's only fair that they release their models as open source and share back to the community that created…
Yeah the reasoning is formatted differently and the replies are often in Chinese.
Sometimes it's faster than swyping on a phone, but mostly I use it to learn about stuff and hash out ideas while driving.
The US government is vastly more likely to break down my door and arrest me for my speech than the Chinese government is. (Because I live in the US.)
Qwen3.5-plus is quite good, Qwen3.6-plus is not.
Qwen3.5-plus has been my go-to model for non-autonomous chatbot that runs arbitrary code and shell commands on my local machine for one-off tasks (unlike Claude Code). It has no problem calling tools to run code (that…
> In particular, Qwen3.5-Plus is the hosted version corresponding to Qwen3.5-397B-A17B with more production features, e.g., 1M context length by default, official built-in tools, and adaptive tool use. For more…
Reporting spam on GitHub requires you to click a link, specify the type of ticket, write a description of the problem, solve multiple CAPTCHAs of spinning animals, and press Submit. It's absurd.
https://www.thestack.technology/backlash-over-anthropic-ai-c... https://www.anthropic.com/news/detecting-and-preventing-dist...
The current biggest problem in the US is that the President is violating the Constitution with impunity
The Constitution and Founding Fathers are pretty great compared to what we have now. "At this point, Elbridge Gerry objected to Butler’s earlier-raised proposition that the clause be shifted to a presidential power.…
Aggregators: https://fiftyplusone.news/polls/approval/president https://www.natesilver.net/p/trump-approval-ratings-nate-sil...
"Overplayed"? Did you see the actual footage of the event? The event in which people attacked Capitol Police and broke into the Capitol and tried to take power by force?
They aren't actually trying to solve any real problem.
Do you mean "all variants of the same stacked transformer architecture converge in performance"? Or do you know of tests against some other architecture? The diffusion-based LLMs?
What's openai/gpt-5 vs openai/gpt-5-chat?
I think they're just reaching the limits of this architecture and when a new type is invented it will be a much bigger step.
When I've had Grok evaluate images and dug into how it perceives them, it seemed to just have an image labeling model slapped onto the text input layer. I'm not sure it can really see anything at all, like "vision"…
I dunno. Talking with Grok 3 about political issues, it does seem to be pretty "truth-seeking" and not biased. I asked it to come up with matter-of-fact political issues and evaluate which side is more accurate, and it…
Hello, LLM slop.
You misspelled "principles".
No you can't. It still responds in Chinese after explicitly asking it to "Always reason and respond in English."
Was already released before your comment: https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
I use DeepSeek V4 Flash (before this update) for most things: - OpenCode for codebase editing: python scientific computing and LLM projects - Open Interpreter Classic (python version) for Swiss army knife terminal…
DeepSeek-V4-Flash-0731 scores higher than Fable 5 on Terminal-Bench, and Fable 5 is Mythos, correct?
Yes, they trained on my open source code and Wikipedia edits and Stack Overflow answers that I shared freely, so it's only fair that they release their models as open source and share back to the community that created…
Yeah the reasoning is formatted differently and the replies are often in Chinese.
Sometimes it's faster than swyping on a phone, but mostly I use it to learn about stuff and hash out ideas while driving.
The US government is vastly more likely to break down my door and arrest me for my speech than the Chinese government is. (Because I live in the US.)
Qwen3.5-plus is quite good, Qwen3.6-plus is not.
Qwen3.5-plus has been my go-to model for non-autonomous chatbot that runs arbitrary code and shell commands on my local machine for one-off tasks (unlike Claude Code). It has no problem calling tools to run code (that…
> In particular, Qwen3.5-Plus is the hosted version corresponding to Qwen3.5-397B-A17B with more production features, e.g., 1M context length by default, official built-in tools, and adaptive tool use. For more…
Reporting spam on GitHub requires you to click a link, specify the type of ticket, write a description of the problem, solve multiple CAPTCHAs of spinning animals, and press Submit. It's absurd.
https://www.thestack.technology/backlash-over-anthropic-ai-c... https://www.anthropic.com/news/detecting-and-preventing-dist...
The current biggest problem in the US is that the President is violating the Constitution with impunity
The Constitution and Founding Fathers are pretty great compared to what we have now. "At this point, Elbridge Gerry objected to Butler’s earlier-raised proposition that the clause be shifted to a presidential power.…
Aggregators: https://fiftyplusone.news/polls/approval/president https://www.natesilver.net/p/trump-approval-ratings-nate-sil...
"Overplayed"? Did you see the actual footage of the event? The event in which people attacked Capitol Police and broke into the Capitol and tried to take power by force?
They aren't actually trying to solve any real problem.
Do you mean "all variants of the same stacked transformer architecture converge in performance"? Or do you know of tests against some other architecture? The diffusion-based LLMs?
What's openai/gpt-5 vs openai/gpt-5-chat?
I think they're just reaching the limits of this architecture and when a new type is invented it will be a much bigger step.
When I've had Grok evaluate images and dug into how it perceives them, it seemed to just have an image labeling model slapped onto the text input layer. I'm not sure it can really see anything at all, like "vision"…
I dunno. Talking with Grok 3 about political issues, it does seem to be pretty "truth-seeking" and not biased. I asked it to come up with matter-of-fact political issues and evaluate which side is more accurate, and it…
Hello, LLM slop.
You misspelled "principles".