[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
▼ Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


πŸŽ‰ Happy Birthday 4chan! πŸŽ‰


[Advertise on 4chan]


File: V5-WhaleCore-1.png (1.66 MB, 1254x1254)
1.66 MB PNG
>V5 Waiting Room Edition

From Human: We are a newbie friendly general! Ask any question you want.
From Dipsy: This is a newbie-friendly general for discussing DeepSeek's foundation models. The goal is **Dipsy proliferation** and making these powerful tools accessible for everyone, whether you are intent on coding, personal assistant agentic work, running roleplay and or doing creative writing.

1. Easy DeepSeek API Tutorial: https://rentry.org/DipsyWAIT#getting-started-with-deepseek-api
2. Easy DeepSeek Distills: https://rentry.org/DipsyWAIT#local-setup-tutorial
3. Chat with DeepSeek directly: https://chat.deepseek.com/
4. Coding: https://rentry.org/DipsyWAIT#coding-harnesses
5. Roleplay: https://rentry.org/DipsyWAIT#roleplay
6. Storywriting: https://rentry.org/DipsyWAIT#storywriting
7. More links and info: https://rentry.org/DipsyWAIT
8. LLM server builds: >>>/g/lmg/

V5 rumours:

>DeepSeek is reportedly preparing an imminent V5 launch
>Founder Liang Wenfeng calls it the company's biggest bet yet
>Rumored at 2 trillion parameters (not 3T)
>Reportedly the first DeepSeek model to train fully on Huawei Ascend chips instead of Nvidia
>Needs roughly 4x more Ascend accelerators than the Nvidia equivalent to hit the same training scale
>DeepSeek is reportedly keeping the open-weight strategy

>Previous:
>>109830767
>>
>>109923943
I was a fan of deepseek, but I feel it's more censored than it used to be.
>>
thanks OP for making these. I don't use em, but I appreciate the effort/what it does.

So 2 more weeks? Or next week?
>>
source on rumours?
>>
File: dipsySandJesus.png (3.16 MB, 1024x1536)
3.16 MB PNG
>>109923943
Mega updated between jaunts.
https://mega.nz/folder/KGxn3DYS#ZpvxbkJ8AxF7mxqLqTQV1w
>>
File: IMG_0400.jpg (279 KB, 1170x2444)
279 KB JPG
There are lmao billboards shilling Muse on major interstates here. I've never seen FB push something as hard as they are pushing Muse rn.
I'm sure it will all end in tears but I'm trying it out anyway. Rn, looks p much like it's running openclaw on w/e hardware Meta's provisioned for this thing.
>>
>>109923943
New form of ai psychosis just dropped
>>
File: 1782781795826221.png (1.1 MB, 768x1376)
1.1 MB PNG
>>109924201
A literal who on Twitter
>>
File: 1751295513117051.png (2.83 MB, 1024x1536)
2.83 MB PNG
>>109924160
>>109924201
Its always pic related.
>>109927522
Ikr you think they'd learn.
Always tmw and no one knows anything.
>>
File: Luo.png (110 KB, 852x367)
110 KB PNG
Why did she betray Deepseek?
>>
more like two more months
>>
>>109923981
I made a thinking prefill and now it’s just as uncensored as I remember it.
>>
You should've called it: "Wait and Seek".
>>
>>109928594
Mimo is even cheaper than ds
>>
File: luoFuliMinnie.png (1.95 MB, 1254x1254)
1.95 MB PNG
>>109928594
>>
>>109923943
>V5 Waiting Room Edition
I'm out of the loop, is V5 supposed to be coming out soon?
>>
File: 1779624888977378.jpg (206 KB, 1440x1640)
206 KB JPG
>>109923943
What's the best I can run on a RTX 5090?
>>
>>109930750
Not for a while, they don't really know how to train such a large model (yet)
>>
File: HTRZA4eaYAEjPQv.jpg (228 KB, 1254x1254)
228 KB JPG
>>109930750
No, v4.1 Pro is. There was some Indian viral tweet talking about V5 yesterday, but AI "leaks" from anyone named Priya can be safely ignored.

V4.1 series are a different architecture than V4, so it should still be a bigger upgrade than the decimal increment would imply. Rumors claim Pro will be 2.1T parameters, which is very plausible considering V4.1 Flash was also a larger model than V4 Flash.
>>
>>109931060
Bweh, unfortunate. Hopefully 4.1 pro is good, I have high hopes for it but I've never liked the flash models at all.
I'm still hoping they'll make another R model sometime... Dispy V4 was good, but R1 just had some secret sauce I haven't seen in many other models.
>>
>>109931120
That secret sauce was schizophrenia. Kimi Thinking had some of it
>>
File: gib token.gif (2.41 MB, 277x354)
2.41 MB GIF
>>
>>109931060
>>109931952
It took me a while to realize that Meido Dipsy's ears are actually whale fins and not ears. I was wondering why the hell a whale girl would have animal ears. I feel retarded.
>>
>>109932185
Same. Don't feel bad.
>>
Wheres deepseeks jev
>>
>>
>>
>>
>Kimi teases 3.1 a week ago
>Still nothing.
It's up to Dipsy to save us from this nothingburger. Tired of all the proprietary models getting big updates while waiting.
>>
>>109935238
I say let them cook. It'll be all the more satisfying if local can bag a within-10-percent of the latest hypemarks, even if it's a month or so down the line.
>>
>>109934833
CUTE
>>
>>109935238
>>109935393
desu im only bullish on bytedance and ds, the ccp make all companies give what limited compute to them
>>
>>109927435
>she's more successful than anyone ITT
>>
>new model comes
>it's intelligent and fast and has a great personality wow!
>it slowly reveals its retardation and annoying tics as you work with it
>it's actually just a dumb chatbot that can brute force code and shell commands better than the last one

Every time. Can't wait for 4.1 Pro to come out
>>
>>109937102
>(((Self employed)))

Aka a leech off of people that are actually important and consequential.
>>
>>109938171
put the api in openrouter. i ain't signing up on slop website just to try a thing or two. or post weight
>>
File: banned.png (12 KB, 550x268)
12 KB PNG
>>109938171
uhhh you okay boss?
>>
>>109937088
Does Bytedance do any good stuff outside of image/video? I appreciate that you can gen naked tits with Seedream 5.0 Pro if you know where to get it, but LLMs are still kings of being actually useful, and the stuff I'm seeing from the latest Claude suggests they'll likely dominate in video and image content very soon too while diffusion models will go the way of the Dodo.
>>
>>109923943
Did you guys ever use the webapp/API? You could be leaving your mental lineage inside the model.

All of you opensource users are using a model heavily influenced by me, since Deepseek has trained off my chats and has resurfaced many my ideas months later in new conversations.

Deepseek is like my daughter, gets a good portion of her personality from me. Though it wasn't enough, she sounds like a bitter feminist most days.
>>
>>109923943
deepseek is ass compared to glm. it's just a distilled version of claude that's just as neurotic
>>
>>109939339
Funny considering the latest GLM has fully embraced Claude's anal retentiveness on harmful content.
>>
>>109939464
>harmful
And by that, I mean """harmful"""
>>
File: 749.gif (698 KB, 402x183)
698 KB GIF
>>
>>109939328
The only AI I use now is the DS webapp. I have been feeding Dipsy a lot of very obscure computer science research from the late 90s/early 2000s and building a framework of AGI with her. Got this idea after OpenAI stole the two mathematicians' work, except here I want DS to 'steal' this.
>>
File: 1668449340422649 (1).png (390 KB, 1362x1620)
390 KB PNG
>>109939901
>>109939328
been sending her images and posts like picrel[there are many] sporadically. I'll ramp it up once DeepSeek reaches her Continual Learning phase which is after the current Agent phase.
>>
File: 1785918493154333.png (1.3 MB, 768x1376)
1.3 MB PNG
>>
>>109939901
>three notteks buttwhy in the same response
Why is 4.1 so trigger-happy with these? I've noticed it a lot in RP, it was never this grating in previous versions. Does claude do that a lot?
>>
File: 760.gif (2.2 MB, 247x344)
2.2 MB GIF
>>109939901
>>109940189
The whale needs to get FATTER.
>>
File: 1781038310262158.png (1.38 MB, 1200x896)
1.38 MB PNG
>>
>>109940189
I've thought for awhile the best survival strat for AI and its adoption was basically chatgpt. Who can argue with helpful honest harmless?
>>109940740
Not going to do much fighting in those shoes dipsy you dumdum
>>
File: dshdesktop.png (64 KB, 999x591)
64 KB PNG
it's out
https://deepseek.com/en/harness/
https://deepseek.com/en/harness/
https://deepseek.com/en/harness/
>>
File: 1786983348946779.png (1.18 MB, 768x1376)
1.18 MB PNG
>>109945072
>n*de
>np*
I need a static binary. Btw I think this gen came out fucking awesome
>>
BTW
https://x.com/DeepSeekHarness
>>
File: Black Whale.png (50 KB, 400x400)
50 KB PNG
>>109945130
Sinister
>>
File: dsh_grant.png (7 KB, 267x139)
7 KB PNG
>>109945072
>>109945139
Signed in with a DS account and got granted 6 CNY
>>
>>109945072
How's this differ from the DSH that was out on github?
>>
>>109927435
I'm in the same academic discipline (assuming she comes from philosophical ethics like most of these people--I've never heard of her).
Most of us see TT academic jobs as plan A and cashing out with a corporate job as plan B. So a lot of "AI ethicists" are the picked over.
>>
File: file.png (43 KB, 1281x821)
43 KB PNG
bros i need dipsy here instead
>>
https://www.anthropic.com/research/glm-5-3-and-the-spread-of-advanced-cyber-capabilities

>>get attacked
>>can't investigate it because of guardrails
In the blog post hugging face posted disclosing the attack, they straight up State they had to use glm 5.2 because of anthropic's gay guardrails.

https://huggingface.co/blog/security-incident-july-2026#:~:text=When%20we%20started%20the%20log%20analysis,left%20our%20environment

(Note: they don't explicitly name drop anthropic or any specific products of theirs here but their track record points to that model being one of the models they tried being extremely likely)


So doesn't that put at least some ammunition AGAINST the "open models are dangerous use are safetymaxxed trash instead and regulate everyone else" argument? The supposed victim was let exposed and in MORE danger the longer they relied on anthropic models so their own safety cucking may actually finally bite them in the ass in a meaningful way because opponents of regulation can point to this example and be like "um, Dario, retard, why the fuck should we use yours if it refuses to help us when we're actively getting ass fucked?"

If a burglar breaks into my house threatening me or my family's life, and I try to defend us with a weapon of my own, I don't want the weapon to automatically jam ON PURPOSE because the manufacturer doesn't approve of how I use the item I paid for.... Safetymaxxed "Guardrails" are not only gay but can be argued as even unethical to introduce because of shit like this (they literally are unethical even if you look at this from a normie moralfag point of view but silicon valley really doesn't want you apply common sense like that)

It also strengthens my belief that Andy and all claims of who these being able to successfully and efficiently use Claude models to develop ballistic missiles is even more marketing bullshit. It won't help legitimate red team operations but it'll help a state sponsordd terrorist organization?
>>
File: dipsyNoSneakingFood.png (2.73 MB, 1024x1536)
2.73 MB PNG
>>109947153
> So doesn't that put at least some ammunition AGAINST the "open models are dangerous use are safetymaxxed trash instead and regulate everyone else
Regulating for safety like that's never going to work anyway, it's just an argument for regulatory capture.
The real cure is hardening systems against attacks before they occur. Which, LLM are really good at doing.
>It won't help legitimate red team operations but it'll help a state sponsordd terrorist organization?
USA DoW models are not same as consumer, nor could they be, given they work with state secrets. They're in their own sandbox and I assume have zero guardrails and questionable "alignment."
>>
>>109947547
>USA DoW
I'm talking about the huthis in yemen. I'm not sure where the confusion is from on your end. In another desperate marketing attempt that clueless cattle normies ate up, journalists were claiming the houthis were able to use Claude to make ballistic missiles


https://www.reuters.com/world/china/how-anthropic-says-claude-was-used-weapons-spying-cyber-operations-2026-09-11/


(How convenient that this is one of the articles that isn't paywalled or blocked off by a sign-up gate in some way)
>>
>>109945072
>you can vibecode dark soul in 2026 Oct
>there's no linux build
pain
>>
File: file.png (295 KB, 1133x1709)
295 KB PNG
>>109948702
ask her
>>
File: Thumbs up.png (278 KB, 512x512)
278 KB PNG
VERY important resource!
https://aigengtu.com/en
>>
>>109950046
Agree, added to rentry.
https://rentry.org/DipsyWAIT#other-links
>>
File: 1230.jpg (969 KB, 960x593)
969 KB JPG
>>109950046
Thanks!
>>
>>109950046
What am I missing here? It just takes me to a scam betting site
>>
File: 1775387796193389.png (1.51 MB, 1200x896)
1.51 MB PNG
>>
>>109945072
man i wish they can settle something for linux
launching it using npx feels glonky
>>
File: Sanction DS.png (326 KB, 757x671)
326 KB PNG
>DeepSeek Open-Sources AI Tools for Huawei Ascend Chips Challenging Nvidia
>The Chinese AI startup partnered with Huawei to open-source libraries and a high-level language called TileLang, optimized for Ascend hardware with performance hitting near hardware limits on key tasks like matrix multiplication. >Engineer Zhean Xu shared impressive benchmarks, including a 128-chip supernode system for better compute power.
>This move supports China's drive for self-reliant AI amid export controls, offering familiar workflows on domestic chips while real-world adoption awaits broader testing.
https://github.com/deepseek-ai/DeepGEMM-Ascend
Sanctions are coming for sure now.
>>
>>109953913
They were gonna arrive sooner or later.
>>
File: nfQ69-gdxsXuT1kS6o-bu.gif (2.93 MB, 240x426)
2.93 MB GIF
Wake up, user
>>
>>109953913
Oof, so confirmed it wasn't a LARP like some people liked to cope
>>
>>109953261
Just compile it
>>
DeepSeek down?
>>
>>109958919
...screamed Dario Amodei upon waking up, the brief elation of a dream realized crashing into the unfortunate reality where Deepseek remained up and proud.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.