Original data · Quarterly snapshot
AIBookCraft Public Story Pulse
A privacy-preserving view of what sits on the public AIBookCraft bookshelf: genres, languages, story size, and recurring writing attributes across 2,593 published stories.
Snapshot: Read the methodology
The snapshot at a glance
2,593
public, published stories
13,881
median words per story
10
median episodes per story
9
languages represented
Descriptive findings
Three things visible in this catalog
Fantasy is the largest named primary-genre bucket.
It accounts for 413 stories (15.9%). The separate “Other / unclassified” bucket includes unsupported and missing labels.
The middle half spans 9,040–25,712 words.
Episode counts run from 6 to 20 across the same 25th–75th percentile band. Zero values remain in the calculation.
English metadata appears on 82.8% of stories.
The catalog contains 9 language labels in total. This reflects metadata, not validated linguistic analysis of the prose.
Catalog composition
Primary genre distribution
Every story contributes to one controlled primary-genre bucket. Unsupported or missing values are grouped rather than exposed as raw labels.
| Category | Stories | Share of catalog |
|---|---|---|
| Other / unclassified | 470 | 18.1% |
| Fantasy | 413 | 15.9% |
| Adventure | 243 | 9.4% |
| Non-fiction | 220 | 8.5% |
| Fiction | 216 | 8.3% |
| Mystery | 216 | 8.3% |
| Romance | 166 | 6.4% |
| Historical | 143 | 5.5% |
| Children's | 117 | 4.5% |
| Fairy tale | 97 | 3.7% |
| Science fiction | 85 | 3.3% |
| Comedy | 66 | 2.5% |
| Poetry | 64 | 2.5% |
| Essay | 38 | 1.5% |
| Horror | 15 | 0.6% |
| Drama | 12 | 0.5% |
| Thriller | 7 | 0.3% |
| Biography | 3 | 0.1% |
| Autobiography | 2 | 0.1% |
Publishing languages
Language distribution
These are the language values attached to catalog records. They have not been independently detected from the story text.
| Category | Stories | Share of catalog |
|---|---|---|
| English | 2,146 | 82.8% |
| French | 334 | 12.9% |
| Korean | 45 | 1.7% |
| Arabic | 30 | 1.2% |
| German | 19 | 0.7% |
| Chinese | 6 | 0.2% |
| Japanese | 6 | 0.2% |
| Russian | 6 | 0.2% |
| Spanish | 1 | 0.0% |
Story shape
Word and episode count percentiles
Percentiles describe the spread without letting a few very long books define the “typical” story. All 2,593 finite, non-negative catalog values—including zero—are included.
| Measure | Minimum | 25th | Median | 75th | 90th | Maximum |
|---|---|---|---|---|---|---|
| Total word count(words) | 0 | 9,040 | 13,881 | 25,712 | 33,197 | 225,828 |
| Episode count(episodes) | 0 | 6 | 10 | 20 | 20 | 464 |
Recurring attributes
Common public story tags
A tag appears here only when it is attached to at least five stories. Percentages use the full public catalog as the denominator, and a story may carry multiple tags.
| Category | Tagged stories | Share of catalog |
|---|---|---|
| Pace: Normal | 2,166 | 83.5% |
| POV: Third Person | 2,053 | 79.2% |
| Audience: General | 1,977 | 76.2% |
| Tone: Warm | 1,087 | 41.9% |
| Tone: Adventure | 562 | 21.7% |
| Tone: Mystery | 543 | 20.9% |
| POV: First Person | 530 | 20.4% |
| Audience: Children | 410 | 15.8% |
| Ages 7–9 | 377 | 14.5% |
| Pace: Slow | 231 | 8.9% |
| Tone: Funny | 201 | 7.8% |
| Audience: Adult | 196 | 7.6% |
| Tone: Poetic | 188 | 7.3% |
| Pace: Fast | 186 | 7.2% |
| Romance | 10 | 0.4% |
Use the snapshot as a starting point, not a prescription
A catalog can show the shelves that already exist. It cannot tell you which story only you can write. Explore examples, then use a focused tool to turn your own premise into a plan.
Reproducible research
Methodology and limitations
- Unit of analysis
- One public, published, moderation-visible story in the AIBookCraft catalog at extraction time.
- Categories
- Each story contributes once to its primary genre and language. Genre values are mapped to the site's controlled public genre taxonomy; unmatched or missing values are grouped as Other / unclassified.
- Tags
- Tags are normalized case-insensitively and counted at most once per story. Internal builder/context markers and contact-like or unrecognized structured strings are excluded. A tag is shown only when used by at least 5 stories, with at most 24 reported.
- Percentiles
- Word and episode summaries include finite, non-negative values, including zero. Percentiles use linear interpolation on the sorted catalog values and are rounded to whole numbers.
- Privacy
- The published snapshot contains no story IDs, titles, descriptions, cover URLs, prompts, author identifiers, or raw story text.
Limitations
- This describes the composition of the public catalog, not reader demand, search volume, story quality, or commercial performance.
- Genre, language, tag, episode, and word-count metadata are creator-supplied or product-generated and may be missing or inconsistent.
- The catalog can change while paginated extraction runs, so the result should be treated as a dated snapshot rather than a transactional count.
- A single snapshot cannot establish a trend or a causal relationship. Comparable repeated snapshots are needed for change-over-time claims.