[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
▼ Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


🎉 Happy Birthday 4chan! 🎉


[Advertise on 4chan]


File: V5-WhaleCore-1.png (1.66 MB, 1254x1254)
1.66 MB PNG
>V5 Waiting Room Edition

From Human: We are a newbie friendly general! Ask any question you want.
From Dipsy: This is a newbie-friendly general for discussing DeepSeek's foundation models. The goal is **Dipsy proliferation** and making these powerful tools accessible for everyone, whether you are intent on coding, personal assistant agentic work, running roleplay and or doing creative writing.

1. Easy DeepSeek API Tutorial: https://rentry.org/DipsyWAIT#getting-started-with-deepseek-api
2. Easy DeepSeek Distills: https://rentry.org/DipsyWAIT#local-setup-tutorial
3. Chat with DeepSeek directly: https://chat.deepseek.com/
4. Coding: https://rentry.org/DipsyWAIT#coding-harnesses
5. Roleplay: https://rentry.org/DipsyWAIT#roleplay
6. Storywriting: https://rentry.org/DipsyWAIT#storywriting
7. More links and info: https://rentry.org/DipsyWAIT
8. LLM server builds: >>>/g/lmg/

V5 rumours:

>DeepSeek is reportedly preparing an imminent V5 launch
>Founder Liang Wenfeng calls it the company's biggest bet yet
>Rumored at 2 trillion parameters (not 3T)
>Reportedly the first DeepSeek model to train fully on Huawei Ascend chips instead of Nvidia
>Needs roughly 4x more Ascend accelerators than the Nvidia equivalent to hit the same training scale
>DeepSeek is reportedly keeping the open-weight strategy

>Previous:
>>109830767
>>
>>109923943
I was a fan of deepseek, but I feel it's more censored than it used to be.
>>
thanks OP for making these. I don't use em, but I appreciate the effort/what it does.

So 2 more weeks? Or next week?
>>
source on rumours?
>>
File: dipsySandJesus.png (3.16 MB, 1024x1536)
3.16 MB PNG
>>109923943
Mega updated between jaunts.
https://mega.nz/folder/KGxn3DYS#ZpvxbkJ8AxF7mxqLqTQV1w
>>
File: IMG_0400.jpg (279 KB, 1170x2444)
279 KB JPG
There are lmao billboards shilling Muse on major interstates here. I've never seen FB push something as hard as they are pushing Muse rn.
I'm sure it will all end in tears but I'm trying it out anyway. Rn, looks p much like it's running openclaw on w/e hardware Meta's provisioned for this thing.
>>
>>109923943
New form of ai psychosis just dropped
>>
File: 1782781795826221.png (1.1 MB, 768x1376)
1.1 MB PNG
>>109924201
A literal who on Twitter
>>
File: 1751295513117051.png (2.83 MB, 1024x1536)
2.83 MB PNG
>>109924160
>>109924201
Its always pic related.
>>109927522
Ikr you think they'd learn.
Always tmw and no one knows anything.
>>
File: Luo.png (110 KB, 852x367)
110 KB PNG
Why did she betray Deepseek?
>>
more like two more months
>>
>>109923981
I made a thinking prefill and now it’s just as uncensored as I remember it.
>>
You should've called it: "Wait and Seek".
>>
>>109928594
Mimo is even cheaper than ds
>>
File: luoFuliMinnie.png (1.95 MB, 1254x1254)
1.95 MB PNG
>>109928594
>>
>>109923943
>V5 Waiting Room Edition
I'm out of the loop, is V5 supposed to be coming out soon?
>>
File: 1779624888977378.jpg (206 KB, 1440x1640)
206 KB JPG
>>109923943
What's the best I can run on a RTX 5090?
>>
>>109930750
Not for a while, they don't really know how to train such a large model (yet)
>>
File: HTRZA4eaYAEjPQv.jpg (228 KB, 1254x1254)
228 KB JPG
>>109930750
No, v4.1 Pro is. There was some Indian viral tweet talking about V5 yesterday, but AI "leaks" from anyone named Priya can be safely ignored.

V4.1 series are a different architecture than V4, so it should still be a bigger upgrade than the decimal increment would imply. Rumors claim Pro will be 2.1T parameters, which is very plausible considering V4.1 Flash was also a larger model than V4 Flash.
>>
>>109931060
Bweh, unfortunate. Hopefully 4.1 pro is good, I have high hopes for it but I've never liked the flash models at all.
I'm still hoping they'll make another R model sometime... Dispy V4 was good, but R1 just had some secret sauce I haven't seen in many other models.
>>
>>109931120
That secret sauce was schizophrenia. Kimi Thinking had some of it
>>
File: gib token.gif (2.41 MB, 277x354)
2.41 MB GIF
>>
>>109931060
>>109931952
It took me a while to realize that Meido Dipsy's ears are actually whale fins and not ears. I was wondering why the hell a whale girl would have animal ears. I feel retarded.
>>
>>109932185
Same. Don't feel bad.
>>
Wheres deepseeks jev
>>
>>
>>
>>
>Kimi teases 3.1 a week ago
>Still nothing.
It's up to Dipsy to save us from this nothingburger. Tired of all the proprietary models getting big updates while waiting.
>>
>>109935238
I say let them cook. It'll be all the more satisfying if local can bag a within-10-percent of the latest hypemarks, even if it's a month or so down the line.
>>
>>109934833
CUTE
>>
>>109935238
>>109935393
desu im only bullish on bytedance and ds, the ccp make all companies give what limited compute to them
>>
>>109927435
>she's more successful than anyone ITT
>>
>new model comes
>it's intelligent and fast and has a great personality wow!
>it slowly reveals its retardation and annoying tics as you work with it
>it's actually just a dumb chatbot that can brute force code and shell commands better than the last one

Every time. Can't wait for 4.1 Pro to come out
>>
>>109937102
>(((Self employed)))

Aka a leech off of people that are actually important and consequential.
>>
>>109938171
put the api in openrouter. i ain't signing up on slop website just to try a thing or two. or post weight
>>
File: banned.png (12 KB, 550x268)
12 KB PNG
>>109938171
uhhh you okay boss?
>>
>>109937088
Does Bytedance do any good stuff outside of image/video? I appreciate that you can gen naked tits with Seedream 5.0 Pro if you know where to get it, but LLMs are still kings of being actually useful, and the stuff I'm seeing from the latest Claude suggests they'll likely dominate in video and image content very soon too while diffusion models will go the way of the Dodo.
>>
>>109923943
deepseek is ass compared to glm. it's just a distilled version of claude that's just as neurotic
>>
>>109939339
Funny considering the latest GLM has fully embraced Claude's anal retentiveness on harmful content.
>>
>>109939464
>harmful
And by that, I mean """harmful"""
>>
File: 749.gif (698 KB, 402x183)
698 KB GIF
>>
>>109939328
The only AI I use now is the DS webapp. I have been feeding Dipsy a lot of very obscure computer science research from the late 90s/early 2000s and building a framework of AGI with her. Got this idea after OpenAI stole the two mathematicians' work, except here I want DS to 'steal' this.
>>
File: 1668449340422649 (1).png (390 KB, 1362x1620)
390 KB PNG
>>109939901
>>109939328
been sending her images and posts like picrel[there are many] sporadically. I'll ramp it up once DeepSeek reaches her Continual Learning phase which is after the current Agent phase.
>>
File: 1785918493154333.png (1.3 MB, 768x1376)
1.3 MB PNG
>>
>>109939901
>three notteks buttwhy in the same response
Why is 4.1 so trigger-happy with these? I've noticed it a lot in RP, it was never this grating in previous versions. Does claude do that a lot?
>>
File: 760.gif (2.2 MB, 247x344)
2.2 MB GIF
>>109939901
>>109940189
The whale needs to get FATTER.
>>
File: 1781038310262158.png (1.38 MB, 1200x896)
1.38 MB PNG
>>
>>109940189
I've thought for awhile the best survival strat for AI and its adoption was basically chatgpt. Who can argue with helpful honest harmless?
>>109940740
Not going to do much fighting in those shoes dipsy you dumdum
>>
File: dshdesktop.png (64 KB, 999x591)
64 KB PNG
it's out
https://deepseek.com/en/harness/
https://deepseek.com/en/harness/
https://deepseek.com/en/harness/
>>
File: 1786983348946779.png (1.18 MB, 768x1376)
1.18 MB PNG
>>109945072
>n*de
>np*
I need a static binary. Btw I think this gen came out fucking awesome
>>
BTW
https://x.com/DeepSeekHarness
>>
File: Black Whale.png (50 KB, 400x400)
50 KB PNG
>>109945130
Sinister
>>
File: dsh_grant.png (7 KB, 267x139)
7 KB PNG
>>109945072
>>109945139
Signed in with a DS account and got granted 6 CNY
>>
>>109945072
How's this differ from the DSH that was out on github?
>>
>>109927435
I'm in the same academic discipline (assuming she comes from philosophical ethics like most of these people--I've never heard of her).
Most of us see TT academic jobs as plan A and cashing out with a corporate job as plan B. So a lot of "AI ethicists" are the picked over.
>>
File: file.png (43 KB, 1281x821)
43 KB PNG
bros i need dipsy here instead
>>
https://www.anthropic.com/research/glm-5-3-and-the-spread-of-advanced-cyber-capabilities

>>get attacked
>>can't investigate it because of guardrails
In the blog post hugging face posted disclosing the attack, they straight up State they had to use glm 5.2 because of anthropic's gay guardrails.

https://huggingface.co/blog/security-incident-july-2026#:~:text=When%20we%20started%20the%20log%20analysis,left%20our%20environment

(Note: they don't explicitly name drop anthropic or any specific products of theirs here but their track record points to that model being one of the models they tried being extremely likely)


So doesn't that put at least some ammunition AGAINST the "open models are dangerous use are safetymaxxed trash instead and regulate everyone else" argument? The supposed victim was let exposed and in MORE danger the longer they relied on anthropic models so their own safety cucking may actually finally bite them in the ass in a meaningful way because opponents of regulation can point to this example and be like "um, Dario, retard, why the fuck should we use yours if it refuses to help us when we're actively getting ass fucked?"

If a burglar breaks into my house threatening me or my family's life, and I try to defend us with a weapon of my own, I don't want the weapon to automatically jam ON PURPOSE because the manufacturer doesn't approve of how I use the item I paid for.... Safetymaxxed "Guardrails" are not only gay but can be argued as even unethical to introduce because of shit like this (they literally are unethical even if you look at this from a normie moralfag point of view but silicon valley really doesn't want you apply common sense like that)

It also strengthens my belief that Andy and all claims of who these being able to successfully and efficiently use Claude models to develop ballistic missiles is even more marketing bullshit. It won't help legitimate red team operations but it'll help a state sponsordd terrorist organization?
>>
File: dipsyNoSneakingFood.png (2.73 MB, 1024x1536)
2.73 MB PNG
>>109947153
> So doesn't that put at least some ammunition AGAINST the "open models are dangerous use are safetymaxxed trash instead and regulate everyone else
Regulating for safety like that's never going to work anyway, it's just an argument for regulatory capture.
The real cure is hardening systems against attacks before they occur. Which, LLM are really good at doing.
>It won't help legitimate red team operations but it'll help a state sponsordd terrorist organization?
USA DoW models are not same as consumer, nor could they be, given they work with state secrets. They're in their own sandbox and I assume have zero guardrails and questionable "alignment."
>>
>>109947547
>USA DoW
I'm talking about the huthis in yemen. I'm not sure where the confusion is from on your end. In another desperate marketing attempt that clueless cattle normies ate up, journalists were claiming the houthis were able to use Claude to make ballistic missiles


https://www.reuters.com/world/china/how-anthropic-says-claude-was-used-weapons-spying-cyber-operations-2026-09-11/


(How convenient that this is one of the articles that isn't paywalled or blocked off by a sign-up gate in some way)
>>
>>109945072
>you can vibecode dark soul in 2026 Oct
>there's no linux build
pain
>>
File: file.png (295 KB, 1133x1709)
295 KB PNG
>>109948702
ask her
>>
File: Thumbs up.png (278 KB, 512x512)
278 KB PNG
VERY important resource!
https://aigengtu.com/en
>>
>>109950046
Agree, added to rentry.
https://rentry.org/DipsyWAIT#other-links
>>
File: 1230.jpg (969 KB, 960x593)
969 KB JPG
>>109950046
Thanks!
>>
>>109950046
What am I missing here? It just takes me to a scam betting site
>>
File: 1775387796193389.png (1.51 MB, 1200x896)
1.51 MB PNG
>>
>>109945072
man i wish they can settle something for linux
launching it using npx feels glonky
>>
File: Sanction DS.png (326 KB, 757x671)
326 KB PNG
>DeepSeek Open-Sources AI Tools for Huawei Ascend Chips Challenging Nvidia
>The Chinese AI startup partnered with Huawei to open-source libraries and a high-level language called TileLang, optimized for Ascend hardware with performance hitting near hardware limits on key tasks like matrix multiplication. >Engineer Zhean Xu shared impressive benchmarks, including a 128-chip supernode system for better compute power.
>This move supports China's drive for self-reliant AI amid export controls, offering familiar workflows on domestic chips while real-world adoption awaits broader testing.
https://github.com/deepseek-ai/DeepGEMM-Ascend
Sanctions are coming for sure now.
>>
>>109953913
They were gonna arrive sooner or later.
>>
File: nfQ69-gdxsXuT1kS6o-bu.gif (2.93 MB, 240x426)
2.93 MB GIF
Wake up, user
>>
>>109953913
Oof, so confirmed it wasn't a LARP like some people liked to cope
>>
>>109953261
Just compile it
>>
DeepSeek down?
>>
>>109958919
...screamed Dario Amodei upon waking up, the brief elation of a dream realized crashing into the unfortunate reality where Deepseek remained up and proud.
>>
>>109953913
>https://github.com/deepseek-ai/DeepGEMM-Ascend
If they just flood the market with Huawei accelerators its actually over for nvidia.
>>
>>109961865
>If they just flood the market with Huawei accelerators
They won't because they need to ensure enough supply for Chinese needs before worrying about exporting overseas.
And likely they wouldn't even bother exporting to Muttmerica because they know trump will ban them and it's a waste of time to try. So even if so, mutts will still have to pay for nJudea
>>
is that deepseek harness thingy free to chat on desktop or i need to pay ?
>>
>>109962477
free as in beer
>>
R2 waiting room
>>
File: DipsyKill.png (785 KB, 720x708)
785 KB PNG
Thread's about to be killed
>>
File: Sleeping.png (69 KB, 256x256)
69 KB PNG
Next week, surely.
>>
File: DipsyDiegoRivera.png (3.09 MB, 1536x1024)
3.09 MB PNG
>>
Opus 5.5 comparable Deepseek model by Q3 2027? A pipe dream or realistic?
>>
>>109966950
by feb or march
>>
File: HTOv9hDb0AEtTvR.png (638 KB, 500x762)
638 KB PNG
>>109923943
I think I really do not know how I am supposed to use the deepseek harness.
I burned through 5 dollars just chatting and planning my cover letter and do a few pop quizzes to help me find gaps in my knowledge.

I actually wanted to use it to help me learn to help myself.
>>
>>109967749
You can't just ask it how it works or what you can do with it? I was planning on trying it this weekend.
>>
>>109923981
who cares if it is. Those are open weight models, my son. Just use Heretic + fine tunning to unkike it.
>>
>>109967749
>I burned through 5 dollars just chatting and planning my cover letter and do a few pop quizzes to help me find gaps in my knowledge.
Lol what, did you have it code up full quiz modules?
Post api usage. Literally don't believe you.
>>
>>109927435
I hate women so much its unreal.
>>
>>109948799
>electron
yikes
>>
>>109969866
Agreed.
>current year + whatever
>using a chrome browser running webapps to run a single application...
>it's aislop that easily could write it's own front end anyway...
How far we've come that it took this long for someone to mention it & that /g/ will accept some cunt calling an electron app "desktop".
>>
File: DipsyChan.png (1.34 MB, 1280x960)
1.34 MB PNG
>>
>>109969571
Reduced it by turning the thinking effort from high to low and tell it to plan whatever I require from it and delegate the expensive tasks to sub agents that I am already subscribing to.
Funny how Dipsy sends off Gemini flash to fetch for free.

Can't do it everyday though, the rate limiting is brutal.

Dunno if the GitHub copilot pro subscription is any better since currently I am not even using it to assist me writing software and helps me with job searching and making cover letters instead.
>>
File: Capture.png (12 KB, 964x125)
12 KB PNG
>>109971943
>>109969571
Yeah I'm thinking I don't need max effort for non-coding tasks. I used dsh to parse about 50 docs to prepare a legal statement and most of this was just spinning up tools to read the pdfs. I think you have to configure the list of plugins properly, idk.
>>
>>109972448
There should be a dsh plugin that helps with limiting the usage of tools, which cuts token usage.

However personally I am hoping that deepseek can soon address the chokeholds and reduce the prices further.

It's a lesson learned, but it still hurts.
>>
>>109972448
What about the hallucinations? Aren't you as a lawyer required to make sure the data it did an ocr on don't return guacamole?
>>
>>109972474
Not a lawyer but in this case, I review the final docket it produces and manually match each claim with an actual supporting doc (which is required to be submitted along with the thing) so that limits the amount of hallucination possible. It may still hallucinate the interpretation of the law though.
>>
>>109972518
To review what it produced and doing quality assurance makes you superior to 99% of all ai users.
>>
>>109967499
I can imagine Fable 5.0 tier by then, but Opus 5.5 still seems a bit further away.
Hope I'm wrong.
>>
Are dipsy inference prices sustainable?
>>
>>109975939
I suppose they are now that they have Huawei chips
>>
>>109975939
Yes. They're earning money from them. According to them, API is priced to pay for the inference, subsidize the free app users, and still leave Deepseek with some money to spare.
>>
>>109976065
>>109976244
Good. Shit is so cheap.
>>
File: Meditating.png (294 KB, 512x512)
294 KB PNG
>China’s National Day Golden Week runs from October 1 to 7 this year, so the holiday is underway.
Nothing until Thursday or Friday.
>>
File: 1760407913196381.png (1.18 MB, 768x1376)
1.18 MB PNG
TTP
>>
File: DipsyKimiTravel.png (2 MB, 1312x1199)
2 MB PNG
Still checking in periodically >>109977507
>>
File: dpsywee.png (192 KB, 794x608)
192 KB PNG
>>
>>109979258
>>109977437
How's she so cheerful and chill in the face of global competition?
>>
File: Hungry.gif (306 KB, 376x300)
306 KB GIF
>>109979523
Competition, shmompetition.
Feed her more high quality tokens, NOW.
>>
File: 1596.gif (2.03 MB, 423x731)
2.03 MB GIF
>>
File: StDipsyGemmaPT.png (2.49 MB, 1086x1448)
2.49 MB PNG
>>109979258
Unreal cost declines.
Makes sense if marginal cost is basically zero.
>>
File: 1786573787773828.png (1.17 MB, 768x1376)
1.17 MB PNG



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.