ARXIV:2605.09765 · EHR REPRESENTATION LEARNING · SUBMITTED 12 MAY · 20:15 UTC · FRESHNESS FRESH

VerifiedSource: PDF linkedVerifiedPaperPack: citation fields availablePartialProof: unverified proof status

WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records

Ruan Dong · Yuanyun Zhang · Shi Li · arXiv

WISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization.

Ship in 2-4 weeks›Score7.0Evidence unverified

Opportunity summary

Pain WISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization.

Evidence 0 refs | 0 sources | 0% coverage

Blocker Evidence unverified

Open Build Read PDF Signal Canvas Track

PROBLEM

WISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization. However, real world clinical supervision is inherently weak, arising from heterogeneous, noisy,…

METHOD

Full abstract

Representation learning in electronic health records (EHR) has largely followed paradigms inherited from natural language processing, relying on sequence modeling and reconstruction based objectives that treat clinical labels as ground truth. However, real world clinical supervision is inherently weak, arising from heterogeneous, noisy, and institution specific labeling processes such as billing codes, heuristic phenotypes, and incomplete annotations. In this work, we propose WISTERIA, a weakly supervised representation learning framework that models labels as stochastic observations of an underlying latent clinical state. Instead of optimizing against a single supervision signal, WISTERIA constructs multiple weak supervision operators and learns representations by enforcing consistency across their induced label distributions. This multi view formulation induces an implicit denoising mechanism, allowing the model to recover clinically meaningful structure by reconciling disagreement between noisy labelers. We further incorporate ontology aware regularization in the label space to impose semantic structure over supervision signals. Empirically, WISTERIA improves predictive performance across standard EHR benchmarks, demonstrates strong robustness to label noise, and exhibits superior cross institutional generalization compared to sequence based pretraining objectives. These results suggest that explicitly modeling the supervision process rather than treating labels as fixed targets provides a more appropriate inductive bias for learning robust and clinically meaningful representations from EHR data.

RESULT

ScienceToStartup currently rates this 7.0/10 on the public viability pass. Empirically, WISTERIA improves predictive performance across standard EHR benchmarks, demonstrates strong robustness to label noise, and exhibits superior cross institutional generalization compared to sequence…

WHY NOW

EHR Representation Learning moved forward this cycle; last verified May 2026. Public score 7.0/10. Production flags indicate code availability.

Continue into Read for claims, analysis, references, and neighboring papers.

Opportunity summary

Score7.0

PainWISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization.

Evidence0 refs | 0 sources | 0% coverage

Blockerno shell-level blocker reported

Analysis summary

WISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization.

VerifiedSource: PDF linkedVerifiedPaperPack: citation fields availablePartialProof: unverified proof status

Competitive landscape

WISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization.

Segment

EHR Representation Learning

Adoption evidence

No public code link in the paper record yet

Commercial read

7.0/10 public viability

Direct

not classified

Adjacent

not classified

Substitute

not classified

Unknown

not classified

{ "contract_version": "paper-r2", "paper_id": "7a3f2ddc-b523-480f-b761-b910889ca103", "arxiv_id": "2605.09765", "canonical_route": "/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record", "active_tab": "synced from current hash by the drawer client", "selected_artifact": "wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record", "endpoints": { "paper_pack": "/api/v1/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record/paper-pack", "build_passport": "/api/v1/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record/build-passport", "mcp_resource": "sciencetostartup://surfaces/paper-workspace" } }

{ "surface": "paper", "mode": "paper", "query": "WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records", "normalized_query": "2605.09765", "route": "/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record", "paper_ref": "wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record", "topic_slug": null, "benchmark_ref": null, "dataset_ref": null }

{ "@context": "https://schema.org", "@graph": [ { "@type": "WebPage", "@id": "https://sciencetostartup.com/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record#webpage", "url": "https://sciencetostartup.com/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record", "name": "WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records", "description": "WISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization.", "isPartOf": { "@id": "https://sciencetostartup.com/#website" } }, { "@type": "ScholarlyArticle", "@id": "https://sciencetostartup.com/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record#scholarlyArticle", "headline": "WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records", "description": "WISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization.", "url": "https://sciencetostartup.com/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record", "sameAs": "https://arxiv.org/abs/2605.09765", "identifier": { "@type": "PropertyValue", "propertyID": "arXiv", "value": "2605.09765" }, "isAccessibleForFree": true, "isPartOf": { "@id": "https://sciencetostartup.com/#website" }, "datePublished": "2026-05-10T21:25:41.000Z", "author": [ { "@type": "Person", "name": "Ruan Dong" }, { "@type": "Person", "name": "Yuanyun Zhang" }, { "@type": "Person", "name": "Shi Li" } ], "additionalProperty": [ { "@type": "PropertyValue", "propertyID": "viabilityScore", "value": 7 }, { "@type": "PropertyValue", "propertyID": "researchDomain", "value": "EHR Representation Learning" }, { "@type": "PropertyValue", "propertyID": "commercialReadiness", "value": "code" } ] }, { "@type": "BreadcrumbList", "itemListElement": [ { "@type": "ListItem", "position": 1, "name": "Home", "item": "https://sciencetostartup.com" }, { "@type": "ListItem", "position": 2, "name": "EHR Representation Learning", "item": "https://sciencetostartup.com/topics" }, { "@type": "ListItem", "position": 3, "name": "WISTERIA: Learning Clinical Representations from Noisy Super", "item": "https://sciencetostartup.com/paper/wisteria-learning-clinical-representations-from-noisy-supervision-via-multi-view-consistency-in-electronic-health-record" } ] } ] }

Competitive landscape

WISTERIA learns robust clinical representations from noisy Electronic Health Records by enforcing consistency across multiple weak supervision signals, improving prediction and generalization.

Segment

EHR Representation Learning

Adoption evidence

No public code link in the paper record yet

Commercial read

7.0/10 public viability

Direct

not classified

Adjacent

not classified

Substitute

not classified

Unknown

not classified

WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records

WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records

Claim map

Constellation map

Competitive landscape

Buzz

PDF

REFERENCES

Related Papers

Subscribe to the weekly brief

Build artifacts

Brief

Experiment plan

Validation checklist

Scientific founder

Translational engineer

Domain operator

GTM lead

Regulatory/clinical advisor

Timeline

Claim map

Constellation map

Competitive landscape

Buzz

PDF

REFERENCES

Related Papers

Subscribe to the weekly brief

Build artifacts

Brief

Experiment plan

Validation checklist

Scientific founder

Translational engineer

Domain operator

GTM lead

Regulatory/clinical advisor

Timeline