ARXIV:2604.03180 · TOPIC MODELING WITH LLMS · SUBMITTED 06 APR · 20:16 UTC · FRESHNESS UNKNOWN

VerifiedSource: PDF linkedVerifiedPaperPack: citation fields availablePartialProof: unverified proof status

PRISM: LLM-Guided Semantic Clustering for High-Precision Topics

Connor Douglas · Utkucan Balci · Joseph Aylett-Bullock · arXiv

A topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis.

Blocked on Code›Score5.0Evidence unverified

Opportunity summary

Pain A topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis.

Evidence 0 refs | 0 sources | 0% coverage

Blocker Evidence unverified

Open Build Read PDF Signal Canvas Track

PROBLEM

A topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis. PRISM fine-tunes a sentence encoding model using a sparse set of LLM- provided labels on samples drawn…

METHOD

Full abstract

In this paper, we propose Precision-Informed Semantic Modeling (PRISM), a structured topic modeling framework combining the benefits of rich representations captured by LLMs with the low cost and interpretability of latent semantic clustering methods. PRISM fine-tunes a sentence encoding model using a sparse set of LLM- provided labels on samples drawn from some corpus of interest. We segment this embedding space with thresholded clustering, yielding clusters that separate closely related topics within some narrow domain. Across multiple corpora, PRISM improves topic separability over state-of-the-art local topic models and even over clustering on large, frontier embedding models while requiring only a small number of LLM queries to train. This work contributes to several research streams by providing (i) a student-teacher pipeline to distill sparse LLM supervision into a lightweight model for topic discovery; (ii) an analysis of the efficacy of sampling strategies to improve local geometry for cluster separability; and (iii) an effective approach for web-scale text analysis, enabling researchers and practitioners to track nuanced claims and subtopics online with an interpretable, locally deployable framework.

RESULT

ScienceToStartup currently rates this 5.0/10 on the public viability pass. Across multiple corpora, PRISM improves topic separability over state-of-the-art local topic models and even over clustering on large, frontier embedding models while requiring only…

WHY NOW

Topic Modeling with LLMs moved forward this cycle; last verified April 2026. Public score 5.0/10.

Continue into Read for claims, analysis, references, and neighboring papers.

Opportunity summary

Score5.0

PainA topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis.

Evidence0 refs | 0 sources | 0% coverage

Blockerno shell-level blocker reported

Analysis summary

A topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis.

VerifiedSource: PDF linkedVerifiedPaperPack: citation fields availablePartialProof: unverified proof status

Competitive landscape

A topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis.

Segment

Topic Modeling with LLMs

Adoption evidence

No public code link in the paper record yet

Commercial read

5.0/10 public viability

Direct

not classified

Adjacent

not classified

Substitute

not classified

Unknown

not classified

{ "contract_version": "paper-r2", "paper_id": "fde2d173-3cdd-4fb0-9a9f-9f16789b361e", "arxiv_id": "2604.03180", "canonical_route": "/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics", "active_tab": "synced from current hash by the drawer client", "selected_artifact": "prism-llm-guided-semantic-clustering-for-high-precision-topics", "endpoints": { "paper_pack": "/api/v1/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics/paper-pack", "build_passport": "/api/v1/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics/build-passport", "mcp_resource": "sciencetostartup://surfaces/paper-workspace" } }

{ "surface": "paper", "mode": "paper", "query": "PRISM: LLM-Guided Semantic Clustering for High-Precision Topics", "normalized_query": "2604.03180", "route": "/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics", "paper_ref": "prism-llm-guided-semantic-clustering-for-high-precision-topics", "topic_slug": null, "benchmark_ref": null, "dataset_ref": null }

{ "@context": "https://schema.org", "@graph": [ { "@type": "WebPage", "@id": "https://sciencetostartup.com/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics#webpage", "url": "https://sciencetostartup.com/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics", "name": "PRISM: LLM-Guided Semantic Clustering for High-Precision Topics", "description": "A topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis.", "isPartOf": { "@id": "https://sciencetostartup.com/#website" } }, { "@type": "ScholarlyArticle", "@id": "https://sciencetostartup.com/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics#scholarlyArticle", "headline": "PRISM: LLM-Guided Semantic Clustering for High-Precision Topics", "description": "A topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis.", "url": "https://sciencetostartup.com/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics", "sameAs": "https://arxiv.org/abs/2604.03180", "identifier": { "@type": "PropertyValue", "propertyID": "arXiv", "value": "2604.03180" }, "isAccessibleForFree": true, "isPartOf": { "@id": "https://sciencetostartup.com/#website" }, "datePublished": "2026-04-03T16:56:47.000Z", "author": [ { "@type": "Person", "name": "Connor Douglas" }, { "@type": "Person", "name": "Utkucan Balci" }, { "@type": "Person", "name": "Joseph Aylett-Bullock" } ], "additionalProperty": [ { "@type": "PropertyValue", "propertyID": "viabilityScore", "value": 5 }, { "@type": "PropertyValue", "propertyID": "researchDomain", "value": "Topic Modeling with LLMs" } ] }, { "@type": "BreadcrumbList", "itemListElement": [ { "@type": "ListItem", "position": 1, "name": "Home", "item": "https://sciencetostartup.com" }, { "@type": "ListItem", "position": 2, "name": "Topic Modeling with LLMs", "item": "https://sciencetostartup.com/topics" }, { "@type": "ListItem", "position": 3, "name": "PRISM: LLM-Guided Semantic Clustering for High-Precision Top", "item": "https://sciencetostartup.com/paper/prism-llm-guided-semantic-clustering-for-high-precision-topics" } ] } ] }

Competitive landscape

A topic modeling framework that uses LLMs to guide semantic clustering for precise topic discovery and analysis.

Segment

Topic Modeling with LLMs

Adoption evidence

No public code link in the paper record yet

Commercial read

5.0/10 public viability

Direct

not classified

Adjacent

not classified

Substitute

not classified

Unknown

not classified

PRISM: LLM-Guided Semantic Clustering for High-Precision Topics

PRISM: LLM-Guided Semantic Clustering for High-Precision Topics

Claim map

Constellation map

Competitive landscape

Buzz

PDF

REFERENCES

Related Papers

Subscribe to the weekly brief

Build artifacts

Brief

Experiment plan

Validation checklist

Scientific founder

Translational engineer

Domain operator

GTM lead

Regulatory/clinical advisor

Timeline

Claim map

Constellation map

Competitive landscape

Buzz

PDF

REFERENCES

Related Papers

Subscribe to the weekly brief

Build artifacts

Brief

Experiment plan

Validation checklist

Scientific founder

Translational engineer

Domain operator

GTM lead

Regulatory/clinical advisor

Timeline