![]() |
|
#16
|
|||
|
|||
|
Quote:
Cloud-based AI turns out to be very expensive or runs out of tokens very fast. If possible, could you provide the steps to use this with Qwen 3.8 or Gemma for example through LMStudio? |
|
#17
|
|||
|
|||
|
ChatGPT Work is cloud based AI and is not as expensive or running out of tokens fast. Ive simply dropped an MCP URL into the chat noting it uses SSE and that was enough. But there are connectors to configure and make it more persistent.
|
|
#18
|
||||
|
||||
|
Why LMstudio when you can run Ollama and configure it in a couple of clicks to switch the engine under the Claude app with an Ollama model? Moreover LMStudio is less efficient. The last time I used it had a problem, it tries to load the model entirely in memory.
The way I proposed uses Claude Coworker for the active parts which has longer token windows as far as I remember. And, yes I used it with a 20$ subscription. If you use the real Claude for commercial targets there are several risks: being banned, blacklisted, identified etc. the only safe way is to run on a strong enough NVIDIA a local abliterated model. By the way today OpenAI declared what’s below, about their GPT 6 Astra cybersecurity model on steroids. Its capabilities of doing automatic RCE are definitely super good. "We also tested Astra on SRE-Bench a benchmark that measures whether models can reverse engineer software binaries to understand its core logic without access to raw source code. Astra solved 88.0% of tasks in a single attempt and 99.2% within four attempts, compared with 55.9% and 68.7% for GPT-5.6 Sol, respectively."
__________________
Ŝħůb-Ňìĝùŕřaŧħ ₪) There are only 10 types of people in the world: Those who understand binary, and those who don't http://www.accessroot.com Last edited by Shub-Nigurrath; 09-05-2026 at 05:02. |
| The Following User Says Thank You to Shub-Nigurrath For This Useful Post: | ||
th3tuga (09-05-2026) | ||
|
#19
|
|||
|
|||
|
Fable 5.1 and Astra are both insane. We enter a new era. It is a master AI or be left in the dust situation.
|
|
#20
|
|||
|
|||
|
Quote:
I was considering running llama.cpp in server mode with the models. How would that compare with Ollama? I’m a bit confused: why wouldn’t we get banned if we used Claude Coworker? I had assumed it would still rely on a Claude subscription behind the scenes. Did you mean that we would be using it with our own models instead? I’m not interested in Astra or other frontier models, because I expect they would blacklist or ban us as soon as they suspect we’re trying to reverse-engineer a commercial target. What I’m trying to find is a way to run an abliterated Qwen 3.6 or 3.8 locally with this mcp rather than through a cloud-based service. Need to possibly use it on commercial targets, so I am avoiding Cloud models, at least the ones that ban or blacklist easily. I suspect that squareD may have been asking the same thing yesterday: Code:
https://forum.exetools.com/showpost.php?p=135961&postcount=13 |
|
#21
|
|||
|
|||
|
A real professional sanitizes a target to disguise it as a crackme or educational material. Rookies get banned.
|
|
#22
|
|||
|
|||
|
Thanks, @chants, but I’d really prefer to let @Shub answer the question.
It’s clear you don’t have enough practical experience in this area. Rookie or not, that kind of bluffing won’t work for long with cutting-edge frontier AI models. There are also privacy concerns associated with using cloud-based models. The long-term solution is to use abliterated local models with a good Nvidia card. |
|
#23
|
|||
|
|||
|
No need for cheap unfounded personal attacks about not having enough practical experience or bluffing. I mean you are asking how to integrate an MCP server. Or repurpose an agent harness like Claude Code which isnt that hard to do. Agent harnesses are client software after all. Im happy to provide solutions but such basic and amateurish info would not match most of the audience here. Of course you might be an exception so I apologize, obviously I need to be more sensitive to our mentally handicapped member. Let me know what beginners guides you need.
|
![]() |
| Thread Tools | |
| Display Modes | |
|
|