create-story
For long fiction. Claude keeps the plot and loses the thread: the second paragraph of a scene reads as though the first was a vague memory. This drafts one scene at a time from a scripted pack the drafter cannot widen, fails a scene where a paragraph shares nothing with the one before it, and has a fresh reader check that nobody made up too early.
/plugin install create-story@fledgeling-pluginsNeeds the marketplace added first — how to do that.
Reach for it when
Write long-form fiction whose paragraphs connect: stories, chapters, scenes, novellas, game narrative prose, audiobook scripts, serialised episodes.
Not for
NOT for outlines or synopses alone (write those as documents), not for non-fiction reports (use agent-voice:agent-voice), and not for content in a named person's voice where no story is involved (use that person's content skill).
- agent-voicein this marketplace
What ships with it
- Scripts it runs itself
- 5 reference files
- Measured evals
Say any of this
- loses the thread
- jumps around
- doesn't follow on
- reads like separate paragraphs
- keeps the plot but not the flow
- write the next scene
Taken from the skill’s own trigger description — these are the phrases it listens for. You do not have to match them exactly.
Easily confused with
Ask Claude for a chapter and you get the plot you asked for, told by someone who keeps forgetting what they wrote a paragraph ago. The characters are right. The outline is honoured. And the second paragraph of every scene reads as though it was written after a night's sleep, with the first one a vague memory. Readers feel it as the story losing its train of thought, and it's the reason a model's long fiction is easy to skim and hard to finish.
This skill is a way of writing stories where each paragraph has to take something from the one before it, and where a script checks that it did.
Why it happens
Three deep-research reports were bought on this question (the reports are in docs/deep-research/, all read in full, with 21 claims traced in skills/create-story/references/evidence.md). They agree on four causes, and each one gets a mechanical counter rather than a paragraph of instructions.
The whole manuscript is in the room. When a model can see everything it has written, the early chapters pull on every sentence it writes, and the link to the paragraph it just finished gets diluted. It's the same reason a person can't proofread a book by re-reading the whole thing every time they add a line. Counter: the drafter never sees the manuscript. A script builds its entire input from the story bible, the state at the end of the last scene, the beat card for this one, and the last two paragraphs written. That's the window, and because a script builds it, nobody widens it by hand.
A paragraph is a self-contained idea. Models learned from the web, where paragraphs usually are. So each new paragraph gets treated as a fresh start that fits the theme without following from the last one. Counter: the drafting brief says the rule in one sentence, and the transition audit reads every seam and fails a scene where a paragraph shares nothing with the one before it.
Everyone makes up too early. A study of 1,200 model-written stories found a consistent pull toward reconciliation. A beat that asks for an argument to stay unresolved gets a paragraph where the characters hug. Counter: every scene ends with a written exit state (who's where, holding what, feeling what, and what's still open), and a check fails the scene if a thread the beat kept open has quietly closed.
Thinking in the middle of writing flattens the prose. Reasoning holds the plot together and makes the sentences stiff. Counter: planning happens in one place with thinking on, and drafting happens somewhere else with no instruction to plan, explain or reason at all.
What you do
Say what you want written. Write the next scene. Turn this outline into chapter two. This chapter keeps the plot but reads like every paragraph was written separately; fix the flow.
The first time, you get a story bible and a beat sheet back and nothing else, because prose written from an unapproved outline is prose you'll throw away. Say the beats stand and the scenes get drafted one at a time. Each one is drafted by a fresh subagent that sees only its pack, gated by the audit, read by a second fresh subagent that's never seen the drafting conversation, and closed when two scripts both exit clean. The prose lands on disk; the reply is six lines saying what closed and what didn't.
Whose voice it's in is composed rather than guessed. Name the authors you like and the skill builds the bible's voice from cards: your own voice as the base (Luke's comes from create-luke-content:create-luke-content), then up to three influences, each a card of observable habits (person, tense, how long the sentences run, how much is dialogue, the named moves, what to borrow and what to leave) drawn from interviews, reviews and stylometry rather than from memory of the books. The composed voice declares its bands, and every drafted scene is measured against them.
If you're the author, say so. The skill routes to your voice skill (create-luke-content:create-luke-content for Luke) for the voice rules and runs that skill's lint on every scene as well as its own. It owns what connects to what; your voice skill owns how it reads.
A version you can listen to
Ask for an audiobook version, a read-aloud or an ElevenLabs prompt and you get one file: a synopsis, setup notes for the person doing the pasting, and the speech itself in parts under the paste limit with vocal direction in the model's own tags. A condensed telling is built from the exit states rather than the prose, follows one named route through any forks and names the others as it passes them. A checker fails it on a tag the model doesn't know, an SSML break the model ignores, setup text that would be read aloud, and a part over the limit, and it reports the running time as a range, because a word count is not a duration. The rules come from the ElevenLabs v3 guide as fetched on 5 September 2026, kept in docs/elevenlabs/. It writes the prompt; rendering it spends your credits and stays yours.
What's in the box
story/
bible.md voice, world, cast, and what's excluded
beats.json one card per scene: goal, the one change, how it ends
scenes/<id>.md the prose
state/<id>.json what's true at the end of that scene
packs/<id>.md exactly what the drafter was shown
critique/<id>.md what the fresh reader found
Five scripts, each with a self-test:
context_pack.pybuilds the drafter's whole input and refuses to include more.transition_audit.pyreads each paragraph seam for a pronoun, a connective, a shared name, a shared word or a spoken line, and fails on none. It also fails on a word count outside the beat's band, first-person leaks in a third-person scene, em dashes, and five stock tells (a testament to, tapestry, delve, little did they know, in a world where).story_state.pyvalidates the bible's shapes, and fails a scene whose exit contradicts its beat card.narration_check.pygates a read-aloud script before it goes near a voice model.voice_card.pycomposes the voice from author cards and measures a scene against its bands.
What it won't do
- It won't write from the whole manuscript, even if you ask. If a scene needs more context, the fix is a bible section named on the beat card, not a wider window.
- It won't rewrite prose in the main session. Edits go back to the drafter with the sentences quoted, for the same reason.
- It won't claim a scene is coherent. The audit is a tripwire with named heuristics, and the research is clear that no measure of paragraph-level coherence exists yet. A scene can pass and still drift in a way only a reader catches, and the reply says so when a scene closes with warnings.
- It won't write outlines, synopses or reports. Those are documents, and other skills own them.
Honest limits
Nothing has measured whether this beats asking Claude for the chapter with no skill at all. The scripts are proven to fire on deliberately bad prose and to pass a fixture, the design rests on three reports and 21 traced claims, and the comparison that matters hasn't been run. EVALS.md says exactly what was and wasn't verified, and records that the name and the icon were chosen without the user in the room.
The window width is a judgement. The reports disagree between two paragraphs, 500 words and 2,000 tokens; the default is two paragraphs and it's a flag, not a law.