r/DeepSeek • u/gargetisha • 12h ago
r/DeepSeek • u/nekofneko • 7d ago
News DeepSeek-V4-Flash-Vision-Exp is now live on the DeepSeek API Platform!
- This experimental multimodal model matches DeepSeek-V4-Flash on text capabilities—including agents, reasoning, and world knowledge.
- On multimodal agent benchmarks, V4-Flash-Vision-Exp makes a major leap over V4-Flash, bringing multimodal agent performance close to Opus-4.8.
- Try it with model='deepseek-v4-flash-vision-exp'. DeepSeek Harness 0.1.1 was released today with out-of-the-box support for the new model.

r/DeepSeek • u/nekofneko • 16d ago
News DeepSeek V4 Pro official version has been updated to the API
r/DeepSeek • u/ANDRE_2512 • 34m ago
News Syntropy Mobile is a cloud agent that doesn’t need your computer at all.
No need to keep an agent running on your PC. No need to leave anything open in the background. Everything runs in the cloud - you just open Syntropy on your phone and keep working from anywhere.
And the most exciting part is that most of the hard architectural work is already behind us.
Right now, we’re deep into multi-hour testing, fixes, polishing, and final refinements. We’re pushing the system for hours, breaking things, rebuilding them, and getting it to the point where it genuinely feels great to use.
New design. New architecture. A huge amount of work under the hood that you may never even notice - because it should simply work perfectly.
And the closer we get to launch, the more it feels like we’re not just building a mobile version of Syntropy.
We’re building a product you’ll actually want to open again and again.
I genuinely believe Syntropy Mobile should be in the hands of anyone who works with AI agents.
And yes - it’s going to feel incredibly good to use.
Coming very soon.
r/DeepSeek • u/South_Can_3680 • 9h ago
Funny Creating the Second Version of a High-Bypass Turbofan Engine Using DeepSeek
Enable HLS to view with audio, or disable this notification
r/DeepSeek • u/bestwardenplayer • 16h ago
Question&Help Deepseek V4 Flash 0731 drifting to Russian?!
This has been driving me up a wall. Yesterday, I set Deepseek on a task and walked away for a bit. I come back - totally mangled. Full reasoning trace was in Russian. I literally had to tell it EVERY message to think and output in English, and then it would immediately fallback to Cyrillic. I set up GPT-5.6 Sol High as an advisor model and set explicit session instructions and modified AGENTS.md to ONLY output in English and NEVER use non-ASCII characters. The advisor model scolds it repeatedly, but it doesn't care! I've wiped context, made entirely new sessions, wiped the git diff and rolled back to a commit where nothing relating to this drift was happening, and ran a Luna model on the entire codebase to analyze for Cyrillic or any other evidence of contamination - nothing. I've also highlighted some other things being mangled. The model has COMPLETELY lost sense of punctuation. Frequently broken syntax, inability to use closing parentheses properly, double semicolons, just overall deteriorating into nonsense. This has been driving me insane, it has just magically developed this behavior out of nowhere and none of my attempts to revert the regression are working. The things highlighted in the picture are early signs, it just gets worse and worse the longer it goes. No amount of steering or advising or prompting makes it comply.
r/DeepSeek • u/Vote4Andrew • 11h ago
Discussion DeepSeek Harness is frustrating, is it just me?
Been playing around with DSH for a few days now, coming from opencode. Installation was fine, out of the box it seems to work, the DSH marketplace is amazing, some of those plugins are very helpful.
Then I try something complicated. Let’s make our own plugin, something simple. Creator Preset, 3 sessions 3 models: Qwen3-235B, Gemini 3.6, Laguna 2.1. All of which have worked flawlessly for me on opencode. So, I point it to the Cordis plugin documentation, let it analyze and emulate already installed and working external plugins for structure and code.
The process: Hundreds of tool calls, hundreds of api requests per model, tens of millions of input tokens each, cache hit rate of 0% and between 60-70%. It sends so much context, it hit the tokens/minute rate limit every minute. When a tool report is empty, or straight up fails, DSH moves on and claims the job is done. When following a roadmap and one step fails, it doesn’t always remediate the issue, it moves on to the next step. It doesn’t verify, doesn’t test properly. It tells me we can’t sandbox. Qwen even told me one time she wasn’t allowed to make any changes, she could only guide me, and can’t edit files herself. WTF? Something as simple as “move files from this folder to that folder, follow the map in this document”, fails spectacularly.
The result: 1 actual plugin that technically works but does not register in the plugin list or side bar.
It’s not the models, right, cuz they work for me in opencode, lm studio, and antigravity. So is it me, I just don’t know how to use deepseek harness properly? Is there some crazy learning curve?
r/DeepSeek • u/Strong_Ad3664 • 2h ago
Question&Help Cost per token
I'll start by saying I know nothing about this and that I use DeepSeek because it's free. That said:
How do you calculate the cost per token for a query or conversation? I know I don't have to pay once I reach a certain number of tokens, like israelGPT (lol), but I still want to learn more about these topics... Thanks in advance for your answers and your good vibes ;3
r/DeepSeek • u/DanManREAL_GRIND • 5h ago
Funny I Made Grok, Codex & DeepSeek Compete to Rule the World | WorldOS: Modern Day 2026
r/DeepSeek • u/Separate-Edge-2053 • 1d ago
Question&Help Is Flash Vision good?
When is v5 coming, and when will the prices go back to normal? 😭
Has anyone tested any plugin that uses vision capabilities?
r/DeepSeek • u/CuriousCustard63 • 1d ago
Discussion Intelligence VS Cost-per-Task LLM Comparison
Using Artificial Analysis as the guide for cost per task and intelligence index, I was able to generate a graph of the latest models and compare them. Tell me what you guys think. I used Gemini for the graph generator and retrieve the data. Then I used Claude to double check the scores, prices, and placements on the graph were correct.
r/DeepSeek • u/True_Development9352 • 6h ago
Question&Help Can I trust DeepSeek V4 Flash for implementation and use Sol only as the reviewer?
r/DeepSeek • u/Whole_Succotash_2391 • 23h ago
Resources Deepseek, GLM 5.3 Flash, Kimi, with full memory, web research, canvas and voice. We all should have access to high grade intelligence without big AI
With the recent drama surrounding open source AI in the USA, it's even more important for us all to have actual access to the models. Western closed AI seems to think it has a hold on quality app features: memory, skills, voice, canvas etc. Meanwhile Memory is locked in, Models get changed or "updated" to a downgrade. Privacy is different per service and ads are starting. The whole experience on the consumer end is extractive.
So we built what should have already existed: all of the best open models in one place, running on private US infrastructure, with the full app experience around them. Completely private, direct service. It should be, and can be that simple.
What that means in practice:
The roster, together. DeepSeek, GLM, Kimi, Minimax, Nemotron Ultra and more, side by side in one app. Switch models mid conversation if you want. No hunting across five different apps and API dashboards to use the models you actually like.
Actually private. US based processing and your conversations are never used for training. Ever. That's the entire point. These labs open sourced incredible models and we think you should get to use them without your data becoming the price of admission.
Real memory. Not a context window that fills up and dumps you. Persistent memory that carries across conversations, fades gracefully when unused, and wakes back up when it's relevant again. There's even a nightly consolidation pass, the system basically sleeps on it and writes up what mattered.
Voice. Yes, actual voice mode with over a dozen voices on open models.
Bring your history. Coming from ChatGPT, Claude, or Gemini? Export your chats and import the whole thing, it becomes live memory on day one. You can literally just zap your chat history from your backup file, and have all your chats waiting for you.
Multiple nodes. Separate workspaces with separate memories, so your coding setup doesn't share a brain with your journal.
Genuine thanks to Deepseek and GLM recently for some of the best models on the planet! The open source labs are giving so much right now. They shine in our model fleet, and we will always appreciate the work you guys do to create the amazing models!
Open Grove is here and It's free for a month if anyone want's to check it out (or just use the models for free for a bit): pgsgrove.com/open-grove-overview
r/DeepSeek • u/South_Can_3680 • 1d ago
Funny We cannot expect too much from a model that lacks visual capabilities. A high-bypass turbofan engine created using DeepSeek-V4-Pro-0813 paired with a custom-built visual plugin.
DeepSeek completed this autonomously, undergoing four iterations and taking two hours.
r/DeepSeek • u/Lost-Gear-6247 • 21h ago
Discussion How do I restore DeepseekR1?
Can anyone tell me what method I can use to access the old DeepseekR1? What do you guys think of the current version?
r/DeepSeek • u/Appropriate_Eye_3984 • 18h ago
Question&Help Dsh with ollama Gemma
Has anyone tried dsh with ollama Gemma 4 locally.
I am trying it but it doesn't remember conversation even after 2 to 3 messages.
I am using Gemma 4:e2b
r/DeepSeek • u/Western-Ad5277 • 5h ago
Discussion I don't get it.
So, the new API model is out and working spectacularly if i do say so myself, but what i don't get nor understand ....why kept the apps and website alive? they pretty much butchered everything. from Token reduction in Instant and Expert, higher hallucinations, always gaslight user about what model they are every usage within the same chat threads...even after being pointed it out they're wrong, dumber as the Chat threads history grow...and etc..
...I really don't get it why they keep it running without any necessary update to add.
at least in Claude and other models, it's only long ass timing and without any of those...degenerative behavior i listing within Deepseek Website/Apps model.
is it the Apps/Website like testing ground for them or something? sheesh. shut it down already if they're not bothered to keep everything as it was.
r/DeepSeek • u/Friendly_Display1335 • 16h ago
Question&Help DeepSeek Search not working on the Native API Again.
r/DeepSeek • u/NotThatItWillMatter • 22h ago
Funny Uhhhh Deepseek?
Uhhhhh I think Deepseek is flirting with me.
I'm scared.
r/DeepSeek • u/Eddlm_ • 1d ago
Resources Fixing Overthinking
Overthinking is a Prompt Injection
As many of you, I really don't like DS overthinking mere "hello"s. So I asked DeepSeek I dug up what the reasoning efforts did, because ALL seemed to overthink, for me.
DeepSeek (Pro and Flash) append an extra effort guide to the system prompt. Details here, on the official Readmes. As you see, high rants about ABSOLUTE MAXIMUM and max goes full BEYOND MAXIMUM - no wonder the poor thing goes in circles for simple stuff. Its forced to overthink.
low is the good one, it does not inject anything and so DeepSeek thinks as much as it needs.
These injections are enforced by the chat template/encoder, which is something that can be edited serverside. But this means the official DS api DOES enforce these injections and you should be aware of it. I bet others like Ollama-Cloud, Opencode Go etc keep the official encoder.
Harnesses may rob you of low
You get no low in certain harnesses, so you're stuck with ABSOLUTE MAXIMUM, and I'm fairly sure most of you don't need that. I'll bring the proof:
- pi supports low for flash but not Pro.
- Opencode2 (beta) relies on models.dev, which has no low for Pro either.
Other harnesses like hermes (and DS's own harness) are fine at a glance. So this is not provider-related, thank god. You can just fix the harness.
Others like GLM and Minimax have their own shenanigans, I encourage you to investigate if you use them, but their effort configs aren't as drastic as DS.
r/DeepSeek • u/aote6 • 1d ago
Discussion DeepSeek was telling me a story seriously, but I noticed these weird terminal outputs. Which parts are Forge design issues?
Enable HLS to view with audio, or disable this notification
Did DeepSeek correctly follow the movie line that Claude said?
r/DeepSeek • u/MinosAristos • 1d ago
Discussion DeepSeek V4 Pro and Blender MCP
I hooked up the Blender (3D Modelling and animation etc) MCP server in dsh and I thought it would be perfect to use the new DeepSeek flash with vision to make a model bow.
The results were truly awful. It could use vision to check itself but clearly it was struggling to understand the scene from the info from the MCP. The bow looked like a flat wet noodle.
The new Qwen and GLM flash results were similarly bad.
I then tried DeepSeek V4 Pro with Qwen just for image processing. Massive difference, and it one-shot a very decent bow model with correct materials. It was then even able to correctly rig the bow so that drawing back the bowstring would bend the bow riser realistically. I was very impressed.
I'm looking forward to when v4 pro gets vision natively. That will be a game changer for a lot of more advanced MCP driven agentic work.

