|
arxiv.org
|
Measuring AI "Slop" in Text
|
|
arxiv.org
|
Compared to What? Baselines and Metrics for Counterfactual Prompting
|
|
arxiv.org
|
Do Natural Language Interpretability Methods Convey Privileged Information?
|
|
icml.cc
|
Interpretability Can Be Actionable
|
|
arxiv.org
|
Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence
|
|
arxiv.org
|
Decide less, communicate more: On the construct validity of end-to-end fact-checking in medicine
|
|
arxiv.org
|
Can SAEs reveal and mitigate racial biases of LLMs in healthcare?
|
|
aclanthology.org
|
Standardizing the Measurement of Text Diversity: A Tool and Comparative Analysis
|
|
arxiv.org
|
Do Automatic Factuality Metrics Measure Factuality? A Critical Evaluation
|
|
arxiv.org
|
Learning the Wrong Lessons: Syntactic-Domain Spurious Correlations in Language Models
|
|
arxiv.org
|
Elucidating Mechanisms of Demographic Bias in LLMs for Healthcare
|
|
arxiv.org
|
The Dual-Route Model of Induction
|
|
arxiv.org
|
Who Taught You That? Tracing Teachers in Model Distillation
|
|
arxiv.org
|
Caught in the Web of Words: Do LLMs Fall for Spin in Medical Literature?
|
|
arxiv.org
|
NNsight and NDIF: Democratizing Access to Foundation Model Internals
|
|
arxiv.org
|
Investigating Mysteries of CoT-Augmented Distillation
|
|
arxiv.org
|
Detection and Measurement of Syntactic Templates in Generated Text
|
|
arxiv.org
|
Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs
|
|
arxiv.org
|
Learning from Natural Language Explanations for Generalizable Entity Matching
|
|
arxiv.org
|
Automatically Extracting Numerical Results from Randomized Controlled Trials with Large Language Models
|
|
arxiv.org
|
Infolossqa: Characterizing and recovering information loss in text simplification
|
|
arxiv.org
|
FactPICO: Factuality Evaluation for Plain Language Summarization of Medical Evidence
|
|
arxiv.org
|
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
|
|
arxiv.org
|
Leveraging ChatGPT in Pharmacovigilance Event Extraction: An Empirical Study
|
|
arxiv.org
|
Multilingual Simplification of Medical Texts
|
|
arxiv.org
|
Anticipating Subsequent Tokens from a Single Hidden State
|
|
arxiv.org
|
Appraising the Potential Uses and Harms of LLMs for Medical Systematic Reviews
|
|
arxiv.org
|
Summarizing, Simplifying, and Synthesizing Medical Evidence using GPT-3 (with Varying Success)
|
|
arxiv.org
|
Automatically Summarizing Evidence from Clinical Trials: A Prototype Highlighting Current Challenges
|
|
arxiv.org
|
How Many and Which Training Points Would Need to be Removed to Flip this Prediction?
|
|
arxiv.org
|
RedHOT: A Corpus of Annotated Medical Questions, Experiences, and Claims on Social Media
|
|
arxiv.org
|
That's the Wrong Lung! Evaluating and Improving the Interpretability of Unsupervised Multimodal Encoders for Medical Data
|
|
arxiv.org
|
PHEE: A Dataset for Pharmacovigilance Event Extraction from Text
|
|
arxiv.org
|
Influence Functions for Sequence Tagging Models
|
|
arxiv.org
|
Self-Repetition in Abstractive Neural Summarizers
|
|
arxiv.org
|
Combining Feature and Instance Attribution to Detect Artifacts
|
|
arxiv.org
|
Evaluating Factuality in Text Simplification
|
|
arxiv.org
|
Unsupervised Data Augmentation with Naive Augmentation and without Unlabeled Data
|
|
arxiv.org
|
Disentangling Representations of Text by Masking Transformers
|
|
arxiv.org
|
Biomedical Interpretable Entity Representations
|
|
arxiv.org
|
An Empirical Comparison of Instance Attribution Methods for NLP
|
|
arxiv.org
|
On the Impact of Random Seeds on the Fairness of Clinical Classifiers
|
|
arxiv.org
|
Paragraph-level Simplification of Medical Texts
|
|
arxiv.org
|
Does BERT Pretrained on Clinical Notes Reveal Sensitive Data?
|
|
arxiv.org
|
Understanding Clinical Trial Reports: Extracting Medical Entities and Their Relations
|
|
arxiv.org
|
Generating (Factual?) Narrative Summaries of RCTs: Experiments with Neural Multi-Document Summarization
|
|
arxiv.org
|
Query-Focused EHR Summarization to Aid Imaging Diagnosis
|
|
arxiv.org
|
Semi-Automating Knowledge Base Construction for Cancer Genetics
|
|
arxiv.org
|
ERASER: A Benchmark to Evaluate Rationalized NLP Models
|
|
arxiv.org
|
Learning to Faithfully Rationalize by Construction
|
|
arxiv.org
|
Trialstreamer: Mapping and Browsing Medical Evidence in Real-Time
|
|
arxiv.org
|
Explaining Black Box Predictions and Unveiling Data Artifacts through Influence Functions
|
|
arxiv.org
|
Practical Obstacles to Deploying Active Learning
|
|
arxiv.org
|
Attention is not Explanation
|
|
arxiv.org
|
Predicting Annotation Difficulty to Improve Task Routing and Model Performance for Biomedical Information Extraction
|
|
arxiv.org
|
Inferring Which Medical Treatments Work from Reports of Clinical Trials
|
|
arxiv.org
|
Structured neural topic models for reviews
|
|
arxiv.org
|
Learning to Identify Patients at Risk of Uncontrolled Hypertension Using Electronic Health Records Data
|
|
arxiv.org
|
Learning Disentangled Representations of Texts with Application to Biomedical Abstracts
|
|
arxiv.org
|
Structured Multi-Label Biomedical Text Tagging via Attentive Neural Tree Decoding
|
|
arxiv.org
|
A Corpus with Multi-Level Annotations of Patients, Interventions and Outcomes to Support Language Processing for Medical Literature
|
|
arxiv.org
|
Syntactic Patterns Improve Information Extraction for Medical Search
|
|
arxiv.org
|
A Sensitivity Analysis of (and Practitioners' Guide to) Convolutional Neural Networks for Sentence Classification
|
|
arxiv.org
|
Quantifying Mental Health from Social Media with Neural User Embeddings
|
|
aclweb.org
|
Automating Biomedical Evidence Synthesis: RobotReviewer
|
|
arxiv.org
|
Exploiting Domain Knowledge via Grouped Weight Sharing with Application to Text Categorization
|
|
arxiv.org
|
Retrofitting Concept Vector Representations of Medical Concepts to Improve Estimates of Semantic Similarity and Relatedness
|
|
arxiv.org
|
Active Discriminative Text Representation Learning
|
|
arxiv.org
|
Rationale-Augmented Convolutional Neural Networks for Text Classification
|
|
arxiv.org
|
Modelling Context with User Embeddings for Sarcasm Detection in Social Media
|
|
arxiv.org
|
MGNC-CNN: A Simple Approach to Exploiting Multiple Word Embeddings for Sentence Classification
|
|
arxiv.org
|
Graph-Sparse LDA: a topic model with structured sparsity
|
|
aclweb.org
|
A Generative Joint, Additive, Sequential Model of Topics and Speech Acts in Patient-Doctor Communication
|
|
arxiv.org
|
Do Multi-Document Summarization Models Synthesize?
|
|
arxiv.org
|
Question answering systems for health professionals at the point of care: A systematic review
|
|
gh.bmj.com
|
State of the evidence: a survey of global disparities in clinical trials
|
|
doi.org
|
Interpretability Analysis for Named Entity Recognition to Understand System Predictions and How They Can Improve
|
|
medinform.jmir.org
|
Predicting Unplanned Readmissions Following a Hip or Knee Arthroplasty: Retrospective Observational Study
|
|
doi.org
|
Trialstreamer: A living, automatically updated database of clinical trial reports
|
|
dx.doi.org
|
Machine learning for identifying Randomized Controlled Trials: An evaluation and practitioner's guide
|
|
dx.doi.org
|
An Exploration of Crowdsourcing Citation Screening for Systematic Reviews
|
|
academic.oup.com
|
Identifying Reports of Randomized Controlled Trials (RCTs) via a Hybrid Machine Learning and Crowdsourcing Approach
|
|
jmlr.org
|
Extracting PICO Sentences from Clinical Trial Reports using Supervised Distant Supervision
|
|
jamia.oxfordjournals.org
|
RobotReviewer: Evaluation of a System for Automatically Assessing Bias in Clinical Trials
|
|
aclanthology.org
|
Overview of MSLR2022: A Shared Task on Multi-document Summarization for Literature Reviews
|
|
aclanthology.org
|
Learning to Ask Like a Physician
|
|
arxiv.org
|
Intermediate Entity-based Sparse Interpretable Representation Learning
|
|
arxiv.org
|
What Would it Take to get Biomedical QA Systems into Practice?
|
|
arxiv.org
|
Evidence Inference 2.0: More Data, Better Models
|
|
arxiv.org
|
An Analysis of Attention over Clinical Notes for Predictive Tasks
|
|
arxiv.org
|
Crowdsourcing Information Extraction for Biomedical Systematic Reviews
|