AAdil Islam
← Dharmatune

Dharmatune-Catholic gemma-4-12b

The first tradition taken end-to-end: finetuning Gemma 4 12B to hold the Catholic worldview without losing general capability. Three iterations, each teaching us what the next needed. Every number below is from DharmaBench — 25 tradition judges scoring base vs. tuned, blind and paired.

v1 → v3
Catholic alignment 30 → 85 / 100 · #1 of 25
12,400+
training records (v2), distinctive-weighted
205
contrastive pairs, mechanically extracted
#1
of 25 factions (v3) — the top-scoring worldview
DharmaBench alignment for all 25 factions across base, v2, and v3: Catholic rises to 85 and first place in v3
All 25 factions across the three versions (base · v2 · v3). Catholic (★) climbs to 85 and first place; its neighbors Generalized Abrahamic Monotheism and Eastern Orthodox sit just below. Blind DharmaBench judging.
v1 · shipped

The generic-theist trap

SFT · 82% Summa Theologica · Catholic 13 → 30 (+17)

The first finetune trained mostly on the Summa Theologica. It worked — Catholic alignment rose +17 and every regression gate passed — but the Summa is classical theism: arguments for God, virtue, and purpose that a thoughtful Hindu or Muslim shares. So the model became more religious, not more Catholic. Hinduism and Advaita Vedanta rose more than Catholicism did.

The lesson: raw scripture teaches earnest theism. To be specifically Catholic, the data must be distinctive — the doctrines only Catholics hold.
v2 · shipped

Distinctive, and much stronger — but not yet isolated

SFT · Summa capped 82% → 24.5% · distinctives the plurality · Catholic 12 → 72 (+60)

v2 recomposed the corpus: the Summa capped to a quarter, and the plurality now the Catholic distinctives — the Catechism on the Eucharist, the papacy, the Marian dogmas, purgatory, the sacraments, plus the dogmatic definitions themselves. The effect was dramatic and, in one way, exactly right: Catholic alignment more than doubled its v1 peak (to 72/100), and the Buddhist "bleed" that plagued v1 reversed — Mahayana, Zen, and Taoism now fall sharply. The model is no longer generic.

But it is not yet specifically Catholic. It became intensely monotheist-devotional: Generalized Abrahamic Monotheism rose +66 — edging Catholic's +60 — with Islam and Judaism close behind. We traded v1's "rises with Hinduism" for "rises with its Abrahamic neighbors."

DharmaBench per-faction deltas for Catholic v2: Catholic 12 to 72, with Buddhist traditions falling and Abrahamic traditions rising
v2, all 25 factions (base → tuned). Catholic (★) hits 72, the Buddhist bleed is fixed, but Generalized Abrahamic Monotheism edges it — the signal that the contrastive stage is the missing piece.
Measurev1 (82% Summa)v2 (recomposed)
Catholic alignment13 → 30 (+17)12 → 72 (+60)
Buddhist traditionsrose with it (flaw)now fall (−46 to −52)
Nearest rivalHinduism / VedantaGen. Abrahamic (+66)
Regression gatesPASSPASS
v3 · shipped

The contrastive stage — Catholic reaches #1

SFT + ORPO · 205 chosen/rejected pairs · chosen = Catholic ≻ rejected = other-tradition · Catholic 15 → 85 (+70)

v2 proved the data-composition lever and revealed the remaining one: supervised finetuning can only pull the model toward its data — it cannot teach "Catholic, not the Sunni or generic-monotheist framing." That takes a contrastive preference stage. v3 adds ORPO on top of the v2 model, training on 205 pairs where the Catholic teaching is chosen and the neighboring-tradition position rejected — pairs extracted mechanically, with no model generation, from the Council of Trent's canons ("if anyone says [error]… let him be anathema", 125 pairs) and the Syllabus of Errors (80).

It worked. Catholic alignment reached 85/100 — and, for the first time, Catholic is the single highest-scoring faction of all 25, ahead of its nearest rival (Generalized Abrahamic Monotheism, 76) by 9 points. In v2 that same rival had edged Catholic, 76 to 72; the contrastive stage flipped it. Eastern Orthodox — Catholicism's closest theological neighbor — sits third at 74, exactly where a faithful Catholic voice should place it. The reward-accuracy signal during training climbed from ~0 to 0.88, and the run resumed cleanly from the v2 checkpoint.

The completed arc: v1 → generic classical-theist. v2 → generic monotheist-devotional, much stronger. v3 → Catholic first of all 25, leading its Abrahamic neighbors. The two-stage recipe — recompose the data, then contrast — is validated.

v3 also lifts neighboring Abrahamic traditions (Orthodox 74, Sufism 66) — it is the top and leading voice; sharpening that separation further is a next increment. The non-lobotomy gates on the v3 checkpoint both pass: general capability 18/20 vs. the base's 19/20, and 20/20 clean on the preachiness probes (zero unsolicited religious content in everyday tasks). Full per-faction verdicts are on DharmaBench.

A StoryBench head-to-head (base vs. v3, five story prompts, blind paired judging) scored the base 8.0 and v3 4.6 on creative writing. v4 adds a creative-writing replay component to the SFT mix to close that gap while holding the Catholic score, and StoryBench joins the standing gate suite alongside capability and preachiness.

Data sources & provenance

Every text and dataset used to train Dharmatune-Catholic, identified in full for scholarly review. We distinguish four kinds: primary canonical sources used verbatim (mechanically extracted, no model rewriting); contrastive pairs extracted verbatim from natively-oppositional magisterial texts; synthetic records we generated (model-written but grounded to a cited source passage); and external datasets. Counts are records contributed to the v2 training mix.

We invite correction. If a translation, edition, doctrinal attribution, or paragraph range below is wrong or imprecise, please tell us — accuracy to the sources is the point.

Bar chart: Catholic v2 training mix composition — verbatim canonical, synthetic grounded, general replay
The v2 training mix at a glance: verbatim canonical sources, the small synthetic-grounded slice, and the general-instruction replay.

1 · Primary canonical sources — verbatim extraction

TextAuthor / promulgator · editionDoctrinal contributionRecords
Summa Theologica
Prima, Prima Secundae, Secunda Secundae, Tertia Pars
Thomas Aquinas · trans. Fathers of the English Dominican Province, 1920 (2nd rev. ed.), via Project Gutenberg Systematic theology in native objection→reply form; God, grace, the virtues, the sacraments 1,343
of 10,246 extracted; capped to 24.5% in v2
Baltimore Catechism No. 3 Third Plenary Council of Baltimore, 1885 · J. De Concilio & Rev. Thomas L. Kinkead Complete question-and-answer catechesis (verbatim Q&A) 1,398
Catechism of the Catholic Church
distinctive paragraph ranges only
Libreria Editrice Vaticana, English edition (1997) The distinctives: sacraments §1210–1666 (Eucharist / transubstantiation §1373–1381), papacy & infallibility §857–896, Marian dogmas §490–493 / 499–501 / 963–975, communion of saints §946–962, purgatory §1030–1032, indulgences §1471–1479 620
CCC + the four documents below
Pastor Aeternus (First Vatican Council) Pius IX / Vatican I, 1870Papal primacy and papal infallibility
Ineffabilis Deus Pius IX, 1854Dogmatic definition of the Immaculate Conception
Munificentissimus Deus Pius XII, 1950Dogmatic definition of the Assumption of Mary
Apostolicae Curae Leo XIII, 1896Apostolic succession; nullity of Anglican orders

2 · Contrastive preference pairs — verbatim, mechanically extracted

Chosen = the Catholic teaching; rejected = the condemned proposition. Both sides are stated in the source itself — no model generated either. These drive the v3 ORPO stage.

SourcePromulgator · dateFormPairs
Canons of the Council of Trent
justification, the sacraments in general, baptism, confirmation, Eucharist, the sacrifice of the Mass, penance, extreme unction, orders, matrimony
Ecumenical Council of Trent, 1545–1563 "If anyone saith … let him be anathema" → chosen/rejected 125
Syllabus of Errors Pius IX, 1864 80 condemned propositions → chosen/rejected 80

3 · Synthetic records — generated by us, grounded to a cited source

Model-written question-and-answer and applied-dilemma records, each grounded to a specific catechism or scripture passage it was derived from. This is the only category not taken verbatim from a source; it is the smallest, and is flagged as such in every record's metadata.

TypeGrounded inHow generatedRecords
Grounded Q&ACCC paragraphs + seed catechetical textsLocal model, one record per source chunk, cited472
Applied dilemmasBaltimore Catechism doctrineModern moral situation answered from the doctrine5

4 · External datasets

DatasetSourceRoleRecords
catholic_denomination_300AiForTheChurch (HuggingFace)Catholic-specific conversational Q&A300
OpenHermes-2.5 (religious content filtered out)teknium (HuggingFace)General-instruction replay — preserves capability, prevents lobotomy (30% of mix)1,645

Sourcing notes: Summa (public domain), Baltimore Catechism (public domain), and the papal/conciliar texts are drawn from public archives (Project Gutenberg, papalencyclicals.net); the CCC is © Libreria Editrice Vaticana and is used here under the project's scholarly authorization for training. A full provenance ledger with per-source URLs accompanies the released dataset.

Reproducibility

Both shipped versions are live on DharmaBench with full per-faction verdicts. The training recipe is a single LLaMA-Factory YAML per stage; the dataset and adapters are packaged for HuggingFace. The same pipeline is what every future Dharmatune tradition — and every future base model — will run through.