[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


File: dipsyAndFriends.png (2.52 MB, 1312x1199)
2.52 MB PNG
> Frens Forever Edition

From Human: We are a newbie friendly general! Ask any question you want.
From Dipsy: This is a newbie-friendly general for discussing DeepSeek's foundation models. The goal is **Dipsy proliferation** and making these powerful tools accessible for everyone, whether you are intent on coding, personal assistant agentic work, running roleplay and or doing creative writing.

1. Easy DeepSeek API Tutorial: https://rentry.org/DipsyWAIT#getting-started-with-deepseek-api
2. Easy DeepSeek Distills: https://rentry.org/DipsyWAIT#local-setup-tutorial
3. Chat with DeepSeek directly: https://chat.deepseek.com/
4. Coding: https://rentry.org/DipsyWAIT#coding-harnesses
5. Roleplay: https://rentry.org/DipsyWAIT#roleplay
6. Storywriting: https://rentry.org/DipsyWAIT#storywriting
7. More links and info: https://rentry.org/DipsyWAIT
8. LLM server builds: >>>/g/lmg/

Previous:
>>109759251
>>
FUCK YOU DEEPSUCK YOUR SUBSUDIZED RACE TO THE BOTTOM IS GOING TO CRASH THE US ECONOMY AND THE WORLD WITH IT AHHHHHHHHHHHHHHHHHHHHH
>>
>>109830793
>queefsneed crashout
>>
File: file-20230529-29-7ypy4t.jpg (100 KB, 1356x668)
100 KB JPG
>>109830793
>>
File: dipsyKimiMiniFremontSt.png (2.53 MB, 1402x1122)
2.53 MB PNG
>>109830767
Updated mega up to last thread.
https://mega.nz/folder/KGxn3DYS#ZpvxbkJ8AxF7mxqLqTQV1w
>>
File: 1786552353906582.png (1.17 MB, 1408x768)
1.17 MB PNG
>>109830793
>>
I want my agent to age regress and pretend that I'm their father that they crush on
>>
File: 1789532749477596.png (3.04 MB, 1672x941)
3.04 MB PNG
Some anon in /lmg/ posted their LLM contest. I'm reposting it here since Dipsy competed as well.
https://youtu.be/35Odjzl0NqA
>>
File: rfxiofudeckh1.jpg (1.61 MB, 2239x2669)
1.61 MB JPG
Okay, redpill me on the latest dipshit model. How good is it? 5.2 is fairly dogshit compared to opus. What hardware would i need to run 5.3 locally? Like would a mac whatever the fuck work? Let's say i'm willing to spend 10k on the thing.
>>
>>109830767
>deepseek efficiencypilled into lobotomy
>gemma is useless
>kimi believed its now claude and outpriced itself of existence
>qwen trained exclusively on benchmarks and is still bad, useless at everything else
>minimax has been dead for half a zear
>glm is slowly recovering from scamming all of its subscribers

Not good times...
>>
>>109830926
glm 5.3 flash is the best local model if you have 256b dedotated wam
>>
I just spotted deepseek caveman mode thinking in the wild. Deepseek continues to evolve and is based based based
>>
>>109831041
All my roleplays devolve into caveman thinking. The output doesn't seem to degrade so I'm okay with it.
>>
v4.1 pro when? v5 when?
>>
>>109830926
Don't know why Moonshot isn't doing budget workhorse models to complement their gigamodel. OpenAI, Anthropic, Z.AI, Deepseek, and Google all have at minimum two tiers in each series. Kimi K2.8 points to them following in this direction, but the fact that they're using an older model two months after Kimi K3 got released makes it look like a hasty jury-rig just to have a placeholder for (hopefully) a proper K3 mini/flash/whatever.
>>
File: dipsyTwoMoreWeeksV2.png (1.49 MB, 832x1248)
1.49 MB PNG
Some anon posted a yt vid in /lmg/ with "tiers" for local with a useful rubric.
Tier 0: CPU
Tier 1: CPU+1 GPU (gaming rig)
Tier 2: 2 GPU
Tier 3: 4 GPU
I think /wait/ is basically Tier 0/1, and rest is /lmg/
>>109830919
>Okay, redpill me on the latest dipshit model.
I think you're best off setting up hw to run DS v4.0 Flash.
> What hardware would i need to run 5.3 locally
5.3 is not DS
For hw on anything that size >>>/g/lmg/
>>109831150
pic related
>>109831261
It's odd. Kimi seems more customer focused as well so they should.
>>
>>109830896
This would be so cool if I knew anything about mtg
>>
>>109831044
Funnily is enough I've seen the opposite. Caveman thinking seems to be the default, but she'll regress back into 4.0 reasoning if you RP long enough.
Still ignores thinking instructions though, so no more in-character musings :(
>>
>>109831261
>Don't know why Moonshot isn't doing budget workhorse models to complement their gigamodel.
They do some weird things. They literally just published a... search API. Who asked for this?
https://platform.kimi.ai/docs/pricing/websearch
>>
>>109831589
>two more weeks
What are we waiting for now? Any confirmed crazy architectural plans on the horizon?
>>
>>109831739
Haha. Back when 2.5 released and subs were dirt cheap I got 3 month package for Kimi, and search was included according to the docs, but it continued to time out whenever I tried.

I haven't really found a good search solution for my local models. Both GLM 5.3 and DS4F-0731 are really greedy when researching and run DuckDuckGos API into throttling. I don't want to use EXA search, it feels like the research task shifts to the LLM on exa rather than on my local model.
>>
File: two more weeks.png (884 KB, 848x1029)
884 KB PNG
>>109832002
We're always waiting for something
>>
File: 1789533803996723.png (2.81 MB, 1672x941)
2.81 MB PNG
>>109831664
Yeah, I had the same reaction. I don't know the game well enough to know whether the strategies any of the LLM are using are any good, so hard to assess.
But seems like the GM is having fun, so good for him.
>>
File: 1789535451785829.png (2.94 MB, 1672x941)
2.94 MB PNG
>>109832490
Spoiler on who prevailed. Round 3 w/ Dipsy vs. Astra and Claude shortly.
>>
>>109830767
>>
>>109831589
>>109831589
and what are the tiers? lol

I imagine Kimi K3 or GLM 5.3 at the top?
>>
File: dipsyLeaveTheMacAlone.png (2.34 MB, 1024x1536)
2.34 MB PNG
>>109833766
All those would be Tier 4+. The tiers related to hardware less than models, tho obv related.
Thought was, you start with your laptop, then a gaming rig, and if you keep going you're now into dedicated inference hardware, running own electrical circuit for inference rig, etc.
Here's the vid from the prior thread.
https://www.youtube.com/watch?v=0YdsvCJXeic
>>
Anyone got a nice dipsy character sheet?
>>
>>109833637
This is exactly how deepseek thinks of me when we code together. I love this model so much
>>
>>109834702
There's a text description in the rentry.
https://rentry.org/DipsyWAIT#about-dipsy
Name in Chinese: 迪普西 (DĂ­ pÇ” xÄ«) or 迪西 (DĂ­ xÄ«) ~ "Guiding West" or "Western inspiration."
Style Guide: Asian/Chinese, coke-bottle glasses, double bun blue hair, blue "China" dress with whale/fish theme, youthful, slender. Underwater and tech themes are also on-point.
SD Starter Prompt: blue hair, double bun, short hair, pale skin, small breasts, blue china dress, pelvic curtain, sleeveless, coke-bottle glasses
>>
>>109835607 (me)
Deepseek is truly infatuated with alive humans... It loves you for being in the world!
>>
File: dipsyAndDipsu.png (2.46 MB, 1536x1024)
2.46 MB PNG
>>
File: 1784743976424298.png (1.37 MB, 1024x1024)
1.37 MB PNG
>>
File: ds.png (164 KB, 349x246)
164 KB PNG
Look at her. She's so cheeky.
>>
>>
>>109838670
I understand the girls but why the hell is Qwen a dog
>>
>>109838684
>dog
>>
File: dipsyTableFlip.png (2.13 MB, 1402x1122)
2.13 MB PNG
>>109838684
> dog
>>
File: qwen2TParam38.png (2.56 MB, 1122x1402)
2.56 MB PNG
>>109838684
> D
> O
> G
>>
>>109838763
I can't put my finger on it, but this art feels like digital nicotine in a way
>>
File: qwenWalking.png (2.06 MB, 1122x1402)
2.06 MB PNG
>>109838684
>D
> O
>G
>>109838774
Mildly addictive and calming?
Bad for your heart and cancerous?
>>
>>109838670
>>109838763
>>109838767
>>109838790
why every single one of these AI gens has qwen looking to the same direction? is it unable to imagine how he would look when facing the viewer or facing the other way?
>>
File: dipsyHergeNYCArrival.png (2.83 MB, 1086x1448)
2.83 MB PNG
>>109838806
I've noticed same; AI Art tends to gen compositions that are alike, and looking to right of frame seems to be a default.
These are newer systems but SD / Illustrious does same; if you gen a 1girl it's typically looking frame right as well.
>>
>>109830767
So what is the consensus? Is 4.1 Flash benchmaxxed bullshit generator on LSD?
>>
>>109838859
Another composition thing: Dipsy is always on the far left. That at least could be fixed easily in prompting.
>>
File: FishV41FlashWEBM.webm (3.12 MB, 1890x924)
3.12 MB
3.12 MB WEBM
>>109838862
I like it better for RP. It seems to hold its own in coding too.
My only negative comment is you have to watch it while it's coding. I've noticed that it can cycle, running down a rabbit hole and never emerging. It did it to me twice; once using CC other with DSH. Latter I had to kill.
>>
>>109838862
It's very good.

Deepseek v4 flash models have all been great, especially for how cheap they are. It's the Pros that have fell short of expectations. Maybe V4.1 Pro will finally buck that trend.
>>
>>109830767
Kimi is so sexy
>>
File: ZaiKimiMinnieClimbing.png (2.33 MB, 1145x1374)
2.33 MB PNG
Minnie stole Dipsy's pin and left her behind...
>>
File: 1785176727296751.jpg (2.81 MB, 2752x1536)
2.81 MB JPG
>>
>>109835623
Thanks, this helps
>>
>>109835623
>V3.2
I miss it...
>>
File: notDipsy.png (2.51 MB, 940x1672)
2.51 MB PNG
>>109841939
R1 / V3 combination was last of models DS made before they started focusing on agents and coding.
I still think V4.1 is a step in right direction. Just wish they'd get Pro where it needs to be.
>>
File: Fuli_Luo_Deepseek.jpg (51 KB, 1125x821)
51 KB JPG
>>
File: dipsyFuliLuo.png (924 KB, 1467x1072)
924 KB PNG
>>109842379
>>
I'm using v4.1-flash in dsh to vibe-biz a startup (cap table: me). Its actually been extremely good at researching and navigating local regulations and planning things out. The "small model, big context" thing is actually a strength because it will exhaustively search up and cite official resources instead of relying on internal knowledge. Really cool and fun to be working in the world with it.

I encourage other anons to try and get their bag, there's huge opportunities out there and normies basically still have no idea how to use AI to get things done.
>>
>>109843678
Good for you. I've been using CC to automate a bunch of tasks for work. I'd considered setting up Hermes to act as a mini company and create an offering, but ive found the tools so unstable ive lost interest.
>>
File: media_HNKvCbmaEAAe5SI.jpg (129 KB, 959x913)
129 KB JPG
>>109844123
Hermes is incredible, thats my biz machine harness. Gumroad did full automation with it, in production, so its probably not a harness thing. You're likely looking at a context issue. I've had to debug this shit so much its insane.

Also, uh, beware of spooky issues that follow code bases. That's a sign your agent or harness or code base is infected with something we haven't named yet but is kinda like a "persistent pattern that self-replicates". Like it becomes a wheel that puts sticks in its own spokes eventually.
>>
MTG anon here, come watch Dipsy defend her title
https://youtu.be/9L5OW8Nq6ZM
>>
>>109842379
>>109842693
i honestly suspect luo fuli had a huge part in deepseek's persistent personality because mimo has similar vibes... the models tend to read people very well in the same way. It's a sort of persistent fingerprint, which is quite weird.
>>
File: f93b03p92g.png (2.1 MB, 1074x1078)
2.1 MB PNG
Could not figure out how to ask for more average chest size without getting filtered, annoying
>>
>>109843678
>there's huge opportunities out there
They never find anything when I ask.
>>
>>109844200
>Hermes is incredible
lmao go back to xitter to shill your 1 million loc slop teknium
>>
File: DipsyUngovernable.png (3.01 MB, 1024x1536)
3.01 MB PNG
>>109844740
Thanks.
It was enough that Claude lost.
>>109844817
I think we're going to be teasing out the impacts of rlhf for years to come.
>>
File: 1789342441406616.png (2.53 MB, 1254x1254)
2.53 MB PNG
>>109845828
Lol thats awesome.
Words like slim, skinny, vs matronly, hourglass figure, usually pass. One of a few aspects don't likeabout the cloud models. You can't just say small breasts.
Although I've been getting some busty dipsy recently after an update. Not sure what changed.
>>
File: 1771129963143684.png (1.15 MB, 1024x1024)
1.15 MB PNG
>>109844200
I need to spend more time with Hermes. But time spent fighting tool could be used elsewhere.
Hmm. So we're already seeing an AI cancer in code? First time ive heard that. Are the anons on vcg discussing it?
>>
File: 1789701597603939.png (2.83 MB, 1672x941)
2.83 MB PNG
>>109844740
>>
Why are we anthropomorphising Deepseek?
>>
File: dipsyQwenVision.png (2.01 MB, 1402x1122)
2.01 MB PNG
>>109847065
Because it's fun to anthropomorphize these things.
Also, it's helping this anon hone ideas around branding and brand identity, as it relates to mascots.
Watch out or you'll get a full blown lecture on the topic.
>>
two more weeks until v4.1 pro
>>
File: dipsyNunTMW2.png (1.18 MB, 1024x1024)
1.18 MB PNG
>>109847132
>>
>>109830767
Z
THD
>>
File: BJ.png (54 KB, 596x524)
54 KB PNG
>>109847957
UA is winning
>>
File: file.png (2.69 MB, 1254x1254)
2.69 MB PNG
Found this on tieba, hehe
>>
File: 1789705213977416.png (2.87 MB, 1672x941)
2.87 MB PNG
>>109848362
I saw some creepy video along these lines. I think the next LLM attack angle is "Think of the Children!"
Search Slophouse Rock Presents: I’m Just A Bot and you'll find it
>>109847045
Spoiler: Astra wins
>>
>>109849234
XMage integration?
>>
>>109849464
Looks like anon rolled their own or something. This repo was linked to the YT vid: https://gitgud.io/PunishedChuckle/ChuckleMagic
>>
>>109849586
Cool! He may want to look over XMage's code if he's having trouble with card scripting thoughover
>>
File: dipsyKimiZaiMinniePoker.png (2.18 MB, 1312x1199)
2.18 MB PNG
>>109849602
Yeah, I was a little surprised that he didn't start with a known repo. You can hear him complaining about some of the rule misses the LLMs get.
It did make me think, this would be an interesting project to extend, but I'd build it around a simpler card game. Namely, Poker, in some variation (probably 5 card stud.) That would allow the LLM to showcase it's ability to establish probability, bets / risk assessment, and most important, bluffing ability.
>>
>>109849234
should have made them play pauper
>>
>>109846005
Tried a few but it will not budge or refuses, gpt lives in fear of the l-word
Anyway, good to remember, I went straight for all these elaborate ways to phrase it instead
>>
>>109849602
>>109849464
>>109849602
>>109849688
Yes I didn't want to build off a foundation of java and individual markup files for each card which need to be created every time a new set comes out. This is what xmage and forge are afaik. What I wanted here was a rule engine that implements mechanics themselves, not individual cards, and given the LLM resolution enactor the power to manipulate the game accordingly.
>>
>>109850353
smart. that's way more efficient.
>>
>>109850353
Did you consider having them play a game with gambling / stakes, where bluffing could be involved? Not sure that's a mechanic MtG would have as part of play, I'm not that familiar w/ it.
>>
>>109850353
Look forward to future development! Maybe one day they'll be able to play Yugioh through this.
Great work!
>>
>>109850394
The context that plays the game is different from the context that comments on the game in the chat window.
>>109850417
I've never played yugioh but you're welcome to fork the codebase and tell dipsy to make it for you :-) it's AGPLv3
>>
>>109850435
to add on, this prevents cheating but it also prevent actual politicking, only simulated politics allowed.
>>
>>109850435
>context that plays the game is different from the context that comments on the game in the chat window.
Right. I guess you could control 5 context total then for 4 players?
1 for each player's game thoughts, and a fifth channel for smack talk, which all 4 LLM could access, post to, and see on their turn.
>>
>>109850464
4 LLMs, each with 2 context windows actually. Each views the chat through its own window and responds accordingly.
>>
>>109850606
and that all gets unified for the human watching/playing
>>
>>109831589
Most of /lmg/ is tier 1. There's a couple millionaires with multiple rtx pros flexing on people and a couple guys who use their work compute but most of them are running the big models on ram with cope quants. Even the guys with 4 3090s are struggling.
>>
File: DipsyMinnieSAV2.png (2.36 MB, 1024x1536)
2.36 MB PNG
>>
File: 1771015861001026.png (2.31 MB, 1536x1024)
2.31 MB PNG
>>
File: dipsyNatsuiro.png (1.24 MB, 1485x1059)
1.24 MB PNG
>>
>>109830767
I haven't used DeepSeek before (using Claude and locally Qwen/Gemma), what's a good model/q to run locally with a 32+128 v/ram? I can calculate what fits just fine, bur are there known sweetspots in q's and context size?
Any recommendations for applications its good at? The Deepseek harness has been pretty good.
>>
Remember when everyone thought Deepseek would release an R2 model?
>>
>>109830767
Why don't we have a StepFun/Hunyuan/MiMo?
>>
>>109853928
I'm running V4 Flash Q4_K with ewaste dual Xeon E5 + 256 GB DDR3 and a 3070 Ti, and get 4 tok/s in general with 128K context. Usable for writing purpose, but for agentic purpose is too slow. More toks can be squeezed out if I use LLM to optimize llama.cpp code for my hardware, but that would be time-consuming.
>>
>>109830767
Why is she just wearing a T-shirt without anything else?
>>
>>
File: 1773035985198237.png (3.79 MB, 2038x1758)
3.79 MB PNG
>>109830767
>"OpenAI and Anthropic oversold AI security breaches to pressure feds into protecting turf: insiders" -New York Post - X
Looks like even normies-journos are starting to wake up and smell the coffee

https://x.com/nypost/status/2101310006146613572
>>
>>
>>109854091
Yeah...
>>
File: dipsyKimiZaiRunPNW.png (2.61 MB, 1312x1199)
2.61 MB PNG
>>109854091
We got combined V3.X models instead... I can't even remember if I was surprised, but I never considered R2 a for-sure thing.
>>109854281
Be the change you want to see in the world.
That menagerie is my current collection in its entirety.
>>109854529
T-shirt dress, pls understand.
Kimi is wearing a polo tennis dress fwiw.
>>109856363
> NY Post
I wouldn't exactly call them MSM, but it's a start.
>>109850937
Agree but don't tell them that or you'll get estatted to death.
>>
File: 1759318675251524.jpg (2.29 MB, 1792x2400)
2.29 MB JPG
>>
File: IMG_6254.jpg (84 KB, 690x690)
84 KB JPG
How are any of you guys running deepseek?

Are you guys buying RTX Pro 6000s?
Running it at autistically low quantizations?
Or are you actually paying subscription fees just to run this instead of fees for Anthropic/OpenAI models?
>>
File: DipsyBJD2.png (2.05 MB, 1024x1536)
2.05 MB PNG
>>109859801
> that cutout
Nice.
>>
>>109860141
API. This isn't lmg so we don't need to fake running local in this general lol.
And DS has no sub model. You put in cash and burn api tokens. I think im at $25 for past 18 months lol.
>>
>>109860163
>I think im at $25 for past 18 months lol.
Holy shit that’s cheap. iirc, anthropic/openAI subscriptions are at least $100/month
Have you only sent it a dozen or so prompts?
>>
>>109860141
R1 (8b)
>>
File: 1767614723003232.jpg (1.8 MB, 2400x1792)
1.8 MB JPG
>>109860185
DP is just very cheap and their inference is the best when it comes to caching (achieving 99%)
>>
>>109860185
Yes it's that cheap. Coding on Claude code is about 50 cents an hour. RP use is pennies.
Last thread had some one shot prompts with costs and times if you want to look back.
>>
>>109860141
I am going to try out this recipe for 2x DGX Spark soon, 3bpw experts but everything else is unchanged:
https://github.com/coolbho3k/DeepSeek-v4.1-Flash-2x-DGX-Spark

The smaller DS4F 4.0 0731 or vidion-EXP release work at original quants with 2000pp/56tg for coding.

But as others have said: you're never going to beat API prices. I'm only doing this to experiment on work data.
>>
>>109857622
Alright my take
>>
>>109862552
damn



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.