What's your context size?

Need to know your context size and the profile job.

I have about 21 profiles.

Default manages all of these profiles :

  • career-guide
  • family-manager
  • health-guide
  • legal-help
  • life-planner
  • money-manager
  • researcher

Director manages all of these profiles :

  • content-marketer
  • creative-designer
  • crm-specialist
  • marketing-analyst
  • marketing-strategist
  • performance-marketer
  • seo-specialist
  • social-media-manager
  • platform-engineer
  • software-engineer
  • technical-reviewer

So I'm not sure which context size I should choose. I'm currently using a 200K context size and I'm using a mix of mimo-2.5, minimax-m3, and deepseek-v4-flash for the profiles. I'm doing compression at 120K context size and I would like to know: what context size are you guys using and for what scenario?

For me Hermes says, "Okay 200k is enough and compression at 120K is more than enough and it's preserving the content and the data fine and everything is fine." I also have "preserve last user message 3" or something like that. .

I need your guys' input: do you use more context and why would you do it? I'm always telling Hermes to use Repos github repo , or local Repo to track and make changes to anything I'm trying to code with Hermes.

So any suggestions?

View Poll

reddit.com
u/x_MASE_x — 1 day ago
▲ 65 r/huggingface+1 crossposts

Goodbye Deepseek on Opencode Go

So I was asking yesterday about what to do after the deepseek price increase and some people still recommended opencode go, and now they officially updated the pricing and also capped the deepseek usage to $15.

Again open for suggestions to replace deepseek for heavy user.

u/x_MASE_x — 3 days ago
▲ 86 r/huggingface+1 crossposts

Hetzner Experiments Platform: FREE AI Models!!

Guys I was just visiting Hetzner and this just poped up, FREE AI models for the testing, have fun guys.

They currently have:

DeepSeek-V4-Flash-0731

GLM-5.2-NVFP4

Kimi-K2.7-Code

Qwen/Qwen3.6-35B-A3B-FP8

Link:
https://experiments.hetzner.com/docs/inference

So not a bad collection we got here, the limits are not great but hey its FREE!!

I'll sure test it out right away.

u/x_MASE_x — 4 days ago

Which plans to use now?

So after deepseek-v4-flash increase I'm a little lost on what to use.

I was using minimax plan first when it was per calls and it was great. Now they changed it and it's not a good option. Not smart enough to justify the low usage and not good quantity to use as a small model.

Then I used codex for a long time and it was kinda working with 2 subscriptions. But with the weekly limit being open to use I quickly used all my tokens fast and ended up with even worse down time.

Then opencode go was used and it was doing great and I liked it because of the new deepseek-v4-flash. I was happy with the old one. But the new one became a primary but now both will be gone and I'll have mimo-2.5 haha.

So my monthly usage is:

Fresh input: 311 M Cached input: 4,523 M Output : 89.4 M Total: 5.024 B

So what's my options besides API pay-as-you-go since it will get out of hand fast. And most of my usage right now is personal stuff so no roi.

Please guys let me know the plans that could work with me. Or what are you using.

reddit.com
u/x_MASE_x — 4 days ago

What's codex estimate usage for gpt-5.5 vs gpt-5.4?

Hey guys I'm using codex but mainly using gpt-5.4 because I think it's saving me some tokens.

But I tried to search around and found no answer of it actually matters.

So does using gpt-5.4 makes the tokens or plan last more than with gpt-5.5?

Also which model are you using with hermes. I can only see gpt-5.5. gpt-5.4. And gpt-5.4 mini.

Some people saying 5.3 codex or something but I can't even pick in my model selection. So what's your model of choice guys?

Also should I keep using codex or use z.ai glm5.2 instead?

If someone had tried them both please tell me if it's a good idea to switch to glm plan or not.

reddit.com
u/x_MASE_x — 1 month ago

How to deal with multiple profiles approvals?

So I have now multiple profiles. Each with a specific tasks and a certain Ai model.

For example I have the project manager with gpt-5.4 and marketing lead with gpt-5.4 and social media manager with minimax-m3 and researcher with deepseek v4 flash.

So the problem now that I'm talking mainly to the pm and instructing him to use kanban to assign other profiles to the tasks.

I have now 2 problems I can't find a solution for them.

1: The approval system for the other profiles. I've set it to smart but they always hit a wall and need approval but nothing gets routed to my project manager and my telegram bot.

Should I set it to yolo or what exactly my option?

2: The other profiles and the project manager never response after the work finish. The other profiles almost all the time never signal that the task or kanban task is done. Finish silently even with me "fixing" it multiple times. It's just random if it does work.

And the project manager the profile I'm talking to. NEVER respond when finished. Even when asking it to monitor the board or the tasks or the entire system. It never come back and say okay I'm done. I have to constantly ask the status and are you done to know the progress.

So am I doing something wrong or what? I know kanban could be viewed from the dashboard but I want the project manager to come back to me. Since looking at the board all the time is the same as asking him for status.

Please if someone has any idea how to solve these problems or what's your setup let me know. Thanks

reddit.com
u/x_MASE_x — 2 months ago

Should I switch from hermes?

So what I'm trying to do is to have a semi automated Marketing Agency. I know it's not perfect now and it won't be for some time but at least I want to free myself from doing many things manually.

I tried openclaw the week or 2 weeks when it was a disaster and the gateway's were slow and %100 cpu usage and not responding to and I faced many thing while hermes worked.

So I need to test whatever that worked and I went with hermes and with time I used openclaw less since I used hermes mainly for sometime and I didn't have time to teach openclaw everything again.

But with time hermes looked like it's not the best option. Many things done with a single profile/agent and the memory got large fast. And my first message was maybe 40k tokens or something.

Then I implemented hindsight and that honestly made things way easier and better and I had some good times working with hermes again. However it was still too bloated for a single profile to do everything.

So I heared about kanban board thing and I did create multiple profiles with good hierarchy with project manager. Marketing lead. Technical lead. Social media manager. Web developer. And many more but I faced the problem of approvals. They were failing without notice. The profiles finishing the tasks without notifying the parent profile or the kanban board that it's done.

It was a disaster for me to always having to ask my agent if the board finished or not.

I've told her to keep watching. To monitor the board. I even had to create another profile/agent to keep monitoring the board and tasks but they still failed.

It was really slow and painful. I still don't know what's the problem. But I gave up and went to less chain of commands and went with marketing lead. Technical lead. Creative lead and tried it but it felt like it was not as rich and as powerful system like it was before.

And now I'm stuck not sure where to go from here. Should I try openclaw again?

Should I create smaller profiles and I manage everything myself or what exactly should I do.

I'm using codex for the top tier profiles. With minimax-m3 for the lower ones and deepseek v4 flash for a few things. So it's not really a bad models problem. I think hermes itself is not a good fit for this task or job.

I really liked openclaw and even used the Claw3D and it was really promising but I gave up on it without trying too much I think.

So right now some people still swear by openclaw and the others by hermes. And I think hermes people like it because it's less annoying to setup and seems to be faster and easier to get it going.

However what I really want is a reliable system. And hermes been reliable till I started using multiple profiles and everything went south and I spent my time fixing issues more than working and using it.

So guys I'm not saying openclaw is bad or hermes is bad. I understand that each system for sure has pros and cons. And I just want to get this multiple profiles/agents going.

So should I try openclaw or try hermes again for this task or what exactly my options?

Im thinking about leaving hermes as my personal assistant and openclaw my agency.

This seems to be my sane and reliable setup. But I'm very open to ideas and suggestions.

Am I doing something wrong or what exactly your experience with both system.

Thanks for reading and sorry for the very long post.

reddit.com
u/x_MASE_x — 2 months ago

Should I switch again?

So what I'm trying to do is to have a semi automated Marketing Agency. I know it's not perfect now and it won't be for some time but at least I want to free myself from doing many things manually.

I tried openclaw the week or 2 weeks when it was a disaster and the gateway's were slow and %100 cpu usage and not responding to and I faced many thing while hermes worked.

So I need to test whatever that worked and I went with hermes and with time I used openclaw less since I used hermes mainly for sometime and I didn't have time to teach openclaw everything again.

But with time hermes looked like it's not the best option. Many things done with a single profile/agent and the memory got large fast. And my first message was maybe 40k tokens or something.

Then I implemented hindsight and that honestly made things way easier and better and I had some good times working with hermes again. However it was still too bloated for a single profile to do everything.

So I heared about kanban board thing and I did create multiple profiles with good hierarchy with project manager. Marketing lead. Technical lead. Social media manager. Web developer. And many more but I faced the problem of approvals. They were failing without notice. The profiles finishing the tasks without notifying the parent profile or the kanban board that it's done.

It was a disaster for me to always having to ask my agent if the board finished or not.

I've told her to keep watching. To monitor the board. I even had to create another profile/agent to keep monitoring the board and tasks but they still failed.

It was really slow and painful. I still don't know what's the problem. But I gave up and went to less chain of commands and went with marketing lead. Technical lead. Creative lead and tried it but it felt like it was not as rich and as powerful system like it was before.

And now I'm stuck not sure where to go from here. Should I try openclaw again?

Should I create smaller profiles and I manage everything myself or what exactly should I do.

I'm using codex for the top tier profiles. With minimax-m3 for the lower ones and deepseek v4 flash for a few things. So it's not really a bad models problem. I think hermes itself is not a good fit for this task or job.

I really liked openclaw and even used the Claw3D and it was really promising but I gave up on it without trying too much I think.

So right now some people still swear by openclaw and the others by hermes. And I think hermes people like it because it's less annoying to setup and seems to be faster and easier to get it going.

However what I really want is a reliable system. And hermes been reliable till I started using multiple profiles and everything went south and I spent my time fixing issues more than working and using it.

So guys I'm not saying openclaw is bad or hermes is bad. I understand that each system for sure has pros and cons. And I just want to get this multiple profiles/agents going.

So should I try openclaw or try hermes again for this task or what exactly my options?

Im thinking about leaving hermes as my personal assistant and openclaw my agency.

This seems to be my sane and reliable setup. But I'm very open to ideas and suggestions.

Am I doing something wrong or what exactly your experience with both system.

Thanks for reading and sorry for the very long post.

reddit.com
u/x_MASE_x — 2 months ago