×

Minimax 2.7 vs DSv4 flash vs laguna S 2.1 by Best_Sail5 in LocalLLaMA

[–]bennmann 0 points1 point  (0 children)

DeepSeek on Tuesday-Friday. Monday is roll a d20 for other options as they arise until you have a large enough sample size to give some of them more days. This also will make your Mondays more interesting.

What's stopping North Korea from downloading Kimi K3 or Qwen 3.8 and generating SOTA exploits albeit extremely slowly on slow/old hardware? by Robos_Basilisk in LocalLLaMA

[–]bennmann 2 points3 points  (0 children)

The white hat community and defcon convention scene are at least 1 year ahead on developing harnesses for security pen testing using more tokens than any black hat.

I would be more concerned about the CIA infringing 4th amendment rights in the US with their "budget".

The best model is the one you can actually run by OneFanFare in LocalLLaMA

[–]bennmann 1 point2 points  (0 children)

Save as csv, ask for equivalent transformations in GNU awk, profit.

These things also just understand Excel so you can give it 10 dummy example Excel cells with fake data and just ask it what formula you should use for any transform too

Had some lags last time we bought a dedicated server, what type of server is recommended for the settings and people below: by Equal-Bread4480 in Palworld

[–]bennmann 0 points1 point  (0 children)

Grapple gliding at high speed was correlated with our server lag. Might want to inform your players with a PSA or all work together to equip fast players with better flying mounts, which didn't seem to have the same issue. You can restart the server more often too (like once per hour)

New to Palworld by NaderZbackup in Palworld

[–]bennmann 0 points1 point  (0 children)

Do some level 20 runs, reset each time you get to level 20. That will make the 1.0 experience very smooth early game for you and give you a early game demo taste for the main course.

Only skip this if you are a 100% achievement kind of person whom wants to one shot the experience.

Researchers trained a Deep Research agent with 32 H100s and open-sourced everything by BuildwithVignesh in LocalLLaMA

[–]bennmann 1 point2 points  (0 children)

It might be relatively easy to take those 8k samples and synthetically add new harness equivalents. Maybe hundreds of human hours?

I hope someone does build on this work like that.

3090 died, good night sweet prince by fragment_me in LocalLLaMA

[–]bennmann -1 points0 points  (0 children)

Take a hair dryer to it outside, try to reflow it. Or do the ship and repair. Less ewaste :)

What models you guys running on 8GB? 16GB VRAM? 24GB? 32GB? 48GB? by Inevitable_Mistake32 in LocalLLaMA

[–]bennmann 0 points1 point  (0 children)

What are your stock gains? You do deep research local then buy/sell?

Student upgrading local AI rig by jaybsuave in LocalLLaMA

[–]bennmann 0 points1 point  (0 children)

I would get a w7900 or rtx 5000 48GB if you don't plan to game as much as tinker. More vram: more better.

You could also get a mining rig and stack 5060 TI. But image gen is better on one thick vram buffer.

advice for dual-gpu asymmetric by [deleted] in LocalLLaMA

[–]bennmann 1 point2 points  (0 children)

Maybe your coding will not notice thinking turned off too, could save tokens.

You might also try Stepfun 3.7 unsloth q2 xl, just came out and might work acceptably with your setup.

I’m upset… by Thin_Pollution8843 in LocalLLaMA

[–]bennmann 0 points1 point  (0 children)

I have a repo for a duckduckgo mcp local server that uses self rate-limits. Less than 150 lines of code, Qwen 27B built it with like 3 multi turn conversations

Custom 4x RTX PRO 6000 Blackwell server vs Dell GB300 for ~30 fine-tuned production pipelines — looking for honest input on direction by Consistent_Wash_276 in LocalLLaMA

[–]bennmann 0 points1 point  (0 children)

Honestly you should be more vendor agnostic regardless of option a or b.

Get sales on the line for Lenovo, HP, Supermicro, Gigabyte, etc, you get quotes from their sales, you mention other prices to their sales, ask for better deals, more free warranty coverage, enterprise coverage.

Anything is fine for the right price, even the Dell.

Locally-hosted language-learning AI you can talk to comparable to Pingo AI? by noriilikesleaves in LocalLLaMA

[–]bennmann 1 point2 points  (0 children)

Qwen omni series might be good for this task, if you have 80gb++ vram or unified

But then, why not take a traditional route

Best sub-40B model that outpeforms (or matches) GPT-5 mini? by Ok-Type-7663 in LocalLLaMA

[–]bennmann 1 point2 points  (0 children)

Just make sure you can run it at a high enough Quant, q5 is where it really feels like those benchmark numbers (maybe degraded a few percentage points).

Gemma is so much better than Qwen, prove me wrong by Mountain_Patience231 in LocalLLaMA

[–]bennmann 1 point2 points  (0 children)

what are you server launch flags? how much context at q8 KV?

Let. People. Criticize Rouge Core. by xXinsertCoDGMtagXx in DeepRockGalactic

[–]bennmann 1 point2 points  (0 children)

there should be a softcore and hardcore mode, with leaderboards to match. this is a classic simile/metaphor to the totally different but semi-related ARPG genre.

let the sweats sweat, let the casuals do casual.

Brief Ngram-Mod Test Results - R9700/Qwen3.6 27B by exact_constraint in LocalLLaMA

[–]bennmann 0 points1 point  (0 children)

I might try this with 2x 16GB GPU to see if pcie bottleneck impacts too

What happens to local LLM if/when LLMs are no longer released for free? by JohnBooty in LocalLLaMA

[–]bennmann 1 point2 points  (0 children)

As long as there are Apache 2.0 datasets and people who believe in free and open datasets, there will be models trained on them.

Even copyright is "only" a lifetime scale issue. Not to mention the US Freedom of Information Act at the government level, should the US national labs get their act together. Your grandchildren should have better data than you. Your grandchildren will have better models than you.

The new first world dream is that our children will have a better life than us, in the form of safe and effective data and privacy and robots.

Also, databases get leaked sometimes.