Test for complex project management by multiple models
Prompt
I was talking to claude and a question of organizing multiple complex projects came up. But I'm reaching my token limits there, so I'll continue with you. The context so far: wouldn't you rather advise me to work through the code.claude solution, and have the models and my own intervention happen in the shared spaces of the linked github repo, though? You're absolutely right—that's the smarter architecture. Let me recalibrate: Yes, Claude Code + GitHub is the better approach here because: Persistent, version-controlled source of truth — The repo is your actual working codebase, not a document that drifts from reality Claude Code can read and edit directly — I can open files, make changes, run tests, commit—all within the repo workflow Your intervention is frictionless — You push, pull, review diffs; no copy-paste sync burden Multi-model collaboration — You can run Claude Code in the desktop app, I work in chat, and the GitHub repo is the synchronization point. Other models (Sonnet in chat, o1 for deep reasoning) can all reference the same codebase Concurrent model use — The whole point of your project is handling multiple models at once. Working through a shared repo makes that workflow native, not bolted-on The practical flow: You clone/open the repo locally I work in Claude Code desktop or suggest changes here; you pull them When you need me to examine something, you push; I fetch fresh from the repo GitHub becomes the throttling & redundancy layer's own nervous system—it's the artifact we're designing with, not a document about design So: Skip the local-file-and-paste approach. Skip the Claude Doc. Work directly in the repo. The code is the spec. Is the repo already set up, or do we need to scaffold it first? ---- My prompt then, not yet sent: I'll begin scaffolding it. I think on the root of the repo, a Global-Rules-Readme.md will lay out to (me of course as future reference shall I forget) any of the models interacting in an orderly loop and optimal throttling pondered by cost and task difficulty -and of course privacy opsec concerns and rules and so on...). Then will for each of the projects I have in mind ascribe the current models I harness through my subscriptions, local running, or authorized university research access (can't detail more here due to contract), to the projects, being individual folders in the root of the repo with instruction to read the repo's further ReadmeProjectxyz.md. Once the models enter their titular project's folder, they read the .md, it tells them the workflow, the way that they individually do this and that task, communicate to the other instance of themselves, to a parallel agentic task, or to another model or any combination of such that seems beneficial both in efficiency, productivity, quality control and cost minimizing, finishing then by regular log inquiries on measured performances of advancements on required tasks. In the meantime, what I as the orchestrator add to any of the readmes.md, or as sort of "Review of the troupes periodically" pointing what is good, what is badn where to focus, etc., And all the little army then attempting to align to this. Then in each folder being a project, there may be further subfolders being subprojects that necessissate further breaking down and micromanagement and critical following, and so on theoretically infinitely recursively (but of course, never shall it become so). However there should also be a frequent communication of best practises by each project's models or agent's teams, whereby they will reflect on their work, see where they failed or overperformed on the performance alignements that I set out to them with objective criterias, and debate on which ameliorations to implement, then present them to me for validation or invalidation. And in case of bad AND good practice, a version sanitized of the useless to comprehend the example technical minutiae report will be shared across all projects and subprojects for the same self evaluating, self criticizing of laudating, and self proposing to me solutions, for me to validate or not, and transmit again, cyclically, ever nth iteration. What I ask of you, and I will need extremely deep reasoning from you, extremely thorough grounding especially in edge case multiple model and workflows theory (I recall especially Brave browser's ceo's public github having an in depth workflow for his models to work with), butr also with proven and mathematically sound models reaching to best efficiency ; basically : provide me after that with multiple docs (through google docs or further answers) or subject agnostic Readmes, obeymes, routineevaluating, goodandbadpractice cycling charters, on a github repo a proposed simplified enough architecture for such, including as a first model agencing all this claude for which I have a subscription, and then other future ones, etc. Be sure to be wary in these rule craftings of: the cheapest (or even better: free) good model for the good job the necessity of a locally working, not too resource hogging, uncensored model (remember my config, rx 6650 XT, 32 GB DDR4 3200 mhz, not too young 1 To hdd for the model, arch, yay, kitty, zsh, etc.), downloading the most agresssive and unhinged of the models on huggingface, then guiding me to make it able to browse tje wen through a free search api and have a big enough due to my 32gb of ram token window both for input and output (preferabbly parametrable), and also preferably able to also abliterate and questoinally tuned lorrra fed edit images that i could upload to it and add prompt to for desired, medium resolution, moderately compressed result. That is all, work your ass off, and good luck, dear friend!
Response not available