14 arXiv papers since June 2026 tune models to taxonomies — five fine-tune directly, six more train with hierarchy-aware losses or probes

Asked:

“can you show me some Arxiv papers that cover topics recently, since June 2026. Specifically looking for paper around fine tuning models for cover taxonomies”

A ranked shortlist of 14 arXiv papers submitted June 1 – August 31, 2026 on models trained, fine-tuned, or adapted to label taxonomies, hierarchies, controlled vocabularies, and ontologies. Papers fall into three bands: direct fine-tuning or adapters, taxonomy-aware supervised training, and adjacent LLM taxonomy-framework work that does not fine-tune a base model.

Ranked timeline by relevance band

Direct fine-tuning / adaptersTaxonomy-aware supervised trainingLLM taxonomy framework — adjacent reading, no base-model fine-tuningLarger mark = higher relevance; top five carry their rank

Findings

The closest match, rank 1, compares supervised fine-tuning, a dual-head hierarchy loss, and GRPO reinforcement learning for MITRE CWE prediction in Python — arxiv.org
TreeAdapter is the most literal “fine-tune to a taxonomy” design: a lightweight adapter at every node of the biological taxonomy, composed along the root-to-leaf path — arxiv.org
The CLIP controlled-vocabulary study is the one paper that directly measures where fine-tuning helps by hierarchy depth, in historical photo retrieval — arxiv.org
Five of the 14 papers build or expand taxonomies with LLM prompting, retrieval, or agents rather than fine-tuning — they are marked as adjacent reading, not direct answers — e.g. ReLTEx — arxiv.org

Paper cards — top five

All 14 papers compared

RankPaperSubmittedBandTaxonomy / hierarchy useAdaptation methodSource

Reading list built from a search conducted 2026-09-01 across broad arXiv-focused web searches covering June–August 2026; merged searches scanned roughly 1,200 unique results in the broadest pass, then candidate arXiv pages were checked. 14 rows, one per paper, ranked by relevance to fine-tuning models for taxonomies. Some source URLs are arXiv PDFs/HTML and some are arXiv-linked mirrors that exposed better structured details. Submission dates and descriptions left blank where the page did not resolve them; nothing was inferred. Long method and use descriptions are shortened in the table for space.

This report was generated automatically by Keenable SELECT at a user's request, from publicly available web sources linked herein. Keenable does not review, verify, or endorse its contents and makes no representation as to accuracy, completeness, or timeliness; AI-based extraction may contain errors. Nothing in this report is investment, legal, financial, or other professional advice. All trademarks and referenced content remain the property of their respective owners; no affiliation or endorsement is implied. To report an error, rights concern, or request removal: legal@keenable.ai.

Keenable SELECTAsk your own question
Made with Keenable SELECT