GPT-6 Astra: Bereinigung von Skills und AGENTS.md in Codex

X-PostJoe DevonSammlung

Joe Devon empfiehlt Entwicklern, ihre AGENTS.md-Dateien und Skills für GPT-6 Astra mit Codex überprüfen zu lassen. Grundlage ist ein Leitfaden von Eric Provencher zur Vermeidung überladener Instruktionen älterer Modellgenerationen.

Das Wichtigste

  1. Ältere Instruktionen für Modelle wie GPT-5.6 Sol oder Luna überfrachten GPT-6 Astra mit unnötigen Vorgaben und verbrauchen wertvollen Kontext.
  2. Zu viele oder zu ausführliche Skill-Beschreibungen führen dazu, dass Codex Texte kürzt oder unpassende Workflows lädt.
  3. Progressive Disclosure wird empfohlen: Skill-Dateien sollten als minimale Router agieren, die erst bei Bedarf auf Detaildokumente verweisen.
  4. GPT-6 Astra führt Tests bereits eigenständig aus; veraltete Vorgaben in AGENTS.md erzwingen oft redundante Testläufe oder unnötige Lesevorgänge vor kleinen Änderungen.
  5. Astra interpretiert Sicherheitsgrenzen sehr strikt und bricht Aufgaben mitunter frühzeitig ab, weshalb klare Kriterien für die Fertigstellung und explizite Freigaben für sichere Workflows nötig sind.
  6. Devon nutzt feste Konventionen für Single Sources of Truth (SSOT), verlangt nummerierte Diffs vor Änderungen und warnt vor Inkompatibilitäten bei gemischten Multi-Modell-Setups.

Warum das relevant ist

Fortschritte bei KI-Modellen machen frühere Prompt-Workarounds überflüssig. Veraltete Vorgaben in Konfigurationsdateien wie AGENTS.md bremsen neuere Agenten aus und verschwenden Kontextfenster.

Einordnung

Der Übergang zu GPT-6 Astra verdeutlicht das Problem der 'Instruction Rot': Was bei älteren Modellen wie GPT-5.6 Sol half, behindert neuere Systeme. Provencher zeigt, dass moderne Agenten weniger kleinteilige Rezepte und stattdessen präzise Kontextbegrenzung sowie klare Abbruchkriterien benötigen. Devon weist zudem auf ein praktisches Problem hin: Wenn Teams mehrere Modelle oder Subagenten parallel einsetzen, kollidieren modellspezifische Optimierungen in zentralen Dateien wie AGENTS.md.

Original-Post

Joe Devon

@joedevon · 5. September 2026

Codex users, do this now for Astra: --- Start Prompt --- "Codex, read @pvncher's article t.co/V1BzZOjcwE then audit all of my skills and AGENTS.md files inside ~/Projects" --- End Prompt --- You're welcome.

2818 Likes26 Antworten8553 Lesezeichen680.812 Aufrufe

Auf X ansehen
Weitere Posts im Thread (10)
  1. @nandanpri @pvncher Haha yeah common. I have strong conventions about which docs are SSOT (single source of truth) for what and to NOT write docs anywhere else to keep files from drifting. I also ask it to tell me if it finds stale instructions . Self improving LlMs are the best LLMs.
  2. @liewcf @pvncher It sure can if you steer it right, @liewcf. My favorite way is you download surf cli by @nicopreme and Chrome Dev or Beta, which will exclusively be your agents' browser. You install the surf chrome extension on Dev/Beta. Then log in to X or whatever you want your browser to
  3. @PamphileRoy @pvncher I'm not sure what you're proposing. Just my perspective, but to me LLMs are magic. It's amazing it works at all. For awhile it will need steering and it's tricky because Astra needs different steering than other models. What could the codex team have done different?
  4. @Secondmindsys @pvncher Not bad. RE: read only part, I've got instructions in my AGENTS.md that requires before editing, it must provide a numbered diff so I can accept or reject the change by number, with a cap for size. For every change it it must delete as much as it added. This forces it to speak
  5. @evoclock @pvncher It is a huge problem that models need different AGENTS instructions. I've got codex and claude $200 plans, and tend to only use the best models, but when they delegate to subagents, I worry about the drift. And it may mess up Fable 5.1 because I wired it up to import AGENTS.md.
  6. @renatiltskin @liewcf @pvncher Here's what I do: t.co/niIYZyOlnC
  7. @evb84 @pvncher Yes. Check out @doodlestein's insane repository of tools, it's very easy to set it up. Offhand I think it was CAUT that tracks usage.
  8. @renatiltskin @ikofranzin @pvncher Yup. There's a lot of tools for that. Here's what I do: t.co/niIYZyOlnC
  9. @Bakaburg1 @pvncher Huge problem. Let me know if you solve it lol. I'm still getting used to astra. It's not working well for me yet. Some combination of hooks and aliases to override default behaviour may be needed. Or more likely you need a default that works well for all agents.
  10. @pvncher This tweet blew up in multiple languages. Kind of a weird experience. Also, the reply counts are not accurate. A ton of replies you can only see by clicking through specific tweets. Many seem to be marked spam when they aren't.
Ausgewählte Antworten (5)
  • @liewcf @joedevon @pvncher You know codex can’t read X article, right?
  • @Secondmindsys @joedevon @pvncher Nice. Add this to the prompt: “Keep the audit read-only. For each proposed change, identify its original purpose, where it applies, and what evidence suggests it’s obsolete. Separate model workarounds from project requirements and permission boundaries. For every workaround
  • @nandanpri @joedevon @pvncher Ran a version of this after the Astra unlock. Pointed Codex at AGENTS.md + skills and it caught two stale path rules I’d been pasting by hand. One prompt beats half an hour of “why is it ignoring my conventions.”
  • @AndrewLachlann @joedevon @pvncher The article literally says Astra does this on its own. Why do you need to prompt it to do this? Way too much of this BS 'prompt this or that' stuff online. It's an LLM, just talk to it.
  • @joko76ers @joedevon @pvncher t.co/yB4MzXIpSw t.co/BF5OJUCGKC

Zusammenfassung von KI erstellt (Gemini 3.8 Flash, 27. September 2026). Sie kann Fehler enthalten – maßgeblich ist die Originalquelle.

Inhaltlich ähnlich, ermittelt über die KI-Suche.

  • X-Post:PostHog

    The Evolving Role of AGENTS.md and CLAUDE.md Context Files

    This discussion explores the effectiveness of using context files like AGENTS.md or CLAUDE.md to guide AI coding assistants. While some developers find them essential for project-specific instructions, others argue they often become technical debt that grows stale and hinders model performance.

    1061Lesezeichen48.044Aufrufe

    KI & AI

  • Video

    Video:CURT

    Systematische Entwicklung von Codex-Skills für KI-Agenten

    Nate Herkelman stellt ein sechsstufiges Framework vor, mit dem wiederholbare Fähigkeiten (Skills) für KI-Agenten in Codex und ähnlichen Systemen entwickelt, validiert und kostenoptimiert werden.

    KI & AI· Anleitung

  • Artikel:Nate

    Agent Skills prüfen: Warum geteilte Workflows oft enttäuschen

    Wer fertige Agent Skills in Tools wie Claude Code oder Codex installiert, übernimmt ungeprüft fremde Qualitätsstandards und Arbeitsweisen. Zu viele Skills führen zudem zu Abstrichen bei der Modellleistung.

    KI & AI· Sammlung

  • X-Post:darkzodchi

    GPT-6 Astra im Produktivbetrieb: Kostenfallen und Optimierung

    Ein Leitfaden von darkzodchi analysiert technische Limits und Preisstrukturen von OpenAIs GPT-6 Astra. Trotz eines Kontextfensters von 1,05 Millionen Tokens führen verdeckte Preissprünge und TPM-Beschränkungen schnell zu unerwarteten Kosten.

    2409Lesezeichen224.201Aufrufe

    KI & AI· Anleitung

  • Video

    Video:AI News & Strategy Daily | Nate B Jones

    Agent Skills: Warum wahllose Installationen schaden und wie man sie richtig baut

    Nate B Jones erklärt, warum das bloße Sammeln vorgefertigter KI-Skills aus dem Internet die Leistung von Agenten verschlechtert. Er beleuchtet die interne Funktionsweise wie Ladereihenfolge und Kontextverbrauch und argumentiert für lesbare, maßgeschneiderte Skills statt unreflektierter Installationen.

    KI & AI· Meinung

  • Repository:coleam00/skills

    Cole's AI Skills: Modulare Fähigkeiten für Coding-Agenten

    Das GitHub-Repository 'coleam00/skills' liefert eine Sammlung von 33 strukturierten Skills für KI-Coding-Agenten wie Claude Code. Anstelle einer riesigen Instruktionsdatei (wie einer unübersichtlichen CLAUDE.md) definiert jeder Skill in Markdown eine konkrete Aufgabe und wird erst bei Bedarf in den Prompt geladen. Das System basiert auf einem standardisierten Ablauf von Vorbereitung über Planung und Implementierung bis hin zur Validierung und Code-Review.

    482SternePython

    KI & AI· Tool

Lassen Sie uns über Ihr Projekt sprechen

Standorte

  • Mattersburg
    Johann Nepomuk Bergerstraße 7/2/14
    7210 Mattersburg, Austria
  • Wien
    Ungargasse 64-66/3/404
    1030 Wien, Austria

Dieser Inhalt wurde teilweise mithilfe von KI erstellt.