PAPER EVIDENCE RANKING

论文证据排行榜

基于PaperMiner证据链动态评分。缺失证据保持未知,不等同于零分;不同任务论文不宜只凭总分直接比较。

综合侧重 科研价值 工程落地 综合影响
单项维度 可信度 可复现性 实验充分度 创新性 工程价值 影响力
评分版本:paper-evidence-v3 最低证据覆盖率:35% 至少3篇论文,取Top 20均值并校正样本量
排名 学校 得分 论文数 覆盖率
1 上海交通大学 Unsupervised Anomaly Detection in Brain MRI via Disentangled Anatomy Learning⋆ · LEIQ-Assessor: Multi-dimensional Quality Assessment of Low-light Enhanced Images via Multi-task Learning · Bench2Drive-VL: Benchmarks for Closed-Loop Autonomous Driving with Vision-Language Models 70.49 274 98%
2 Tsinghua University SpikACom: A Neuromorphic Computing Framework for Green Communications · LEGALONE: A FAMILY OF FOUNDATION MODELS FOR RELIABLE LEGAL REASONING · A Foundational EDM2-Based Generative Model for High-Resolution Synthetic Fetal Ultrasound Imaging from Open Datasets 70.47 321 98%
3 Zhejiang University HGeo-TopoMap: Boosting Topological Mapping with Hierarchical Geometric Priors · DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing · On the Step Length Confounding in LLM Reasoning Data Selection 70.31 246 98%
4 Fudan University SWE-bench Science: Can Coding Agents Resolve Engineering Tasks in Science? · ABC-Bench: Benchmarking Agentic Backend Coding in Real-World Development · Bench2Drive-VL: Benchmarks for Closed-Loop Autonomous Driving with Vision-Language Models 70.15 154 96%
5 The Chinese University of Hong Kong StyleDoctor: Towards Specialist Reward Model for Style-centric Generation Tasks · Approaching Low-Cost Cardiac Intelligence with Semi-Supervised Knowledge Distillation · Metamorphic Testing for Audio Content Moderation Software 70.05 145 95%
6 Peking University DALD-PCAC: Density-Adaptive Learning Descriptor for Point Cloud Lossless Attribute Compression · Crystal structure prediction with nuclear quantum and finite-temperature effects via deep free energy learning · VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding 69.96 227 94%
7 Stanford University The limits of fair medical imaging AI in real-world generalization · What LLMs Think When You Don’t Tell Them What to Think About? · CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation Models 69.91 142 95%
8 Massachusetts Institute of Technology The limits of fair medical imaging AI in real-world generalization · Deterministic access to global viral sequence data enables robust agent-driven scientific discovery · Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering 69.81 138 95%
9 National University of Singapore An Event-Based Opto-Tactile Skin · ALPaCA: Adapting Llama for Pathology Context Analysis to enable slide-level question answering · Certified Program Synthesis with a Multi-Modal Verifier 69.68 147 98%
10 University College Causal-Adversarial Probing of Clinical Covariates for Prostate MRI Grading · A Foundational EDM2-Based Generative Model for High-Resolution Synthetic Fetal Ultrasound Imaging from Open Datasets · INTRODUCTION AND NUMERICAL VALIDATION OF AN OPEN-SOURCE MATLAB PACKAGE FOR QUANTITATIVE ULTRASOUND TOMOGRAPHY VIA RAY-BORN INVERSION 69.63 80 95%
11 Wuhan University C2|Q⟩: A Robust Framework for Bridging Classical and Quantum Software Development — RCR Report · Trust Your Critic: Robust Reward Modeling and Reinforcement Learning for Faithful Image Editing and Generation · CoRE-UIR: Prior-guided common and residual experts for efficient all-in-one remote sensing image restoration 69.62 78 95%
12 University of Science and Technology of China DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing · Towards Expressive and Faithful Audio-to-Image Generation: A Unified Multimodal Dataset and Synthesis Framework · NTIRE 2025 Challenge on UGC Video Enhancement: Methods and Results 69.58 183 94%
13 Nanjing University DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing · VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding · RIVER: A REAL-TIME INTERACTION BENCHMARK FOR VIDEO LLMS 69.4 89 94%
14 Sun Yat-sen University DepthAnything and SAM for UIE: Exploring Large Model Information Contributes to Underwater Image Restoration · DepthAnything and SAM for UIE: Exploring Large Model Information Contributes to Underwater Image Restoration · Comment Traps: How Defective Commented-out Code Augment Defects in AI-Assisted Code Generation 69.38 76 94%
15 University of Chinese Academy of Sciences A Mutil-conditional Diffusion Transformer for Versatile Seismic Wave Generation · CL-VISTA: Benchmarking Continual Learning in Video Large Language Models · MGPC: Multimodal Network for Generalizable Point Cloud Completion With Modality Dropout and Progressive Decoding 69.36 139 94%
16 University of Michigan Beyond Screenshots: Evaluating VLMs’ Understanding of UI Animations · EXP-Bench: Can AI Conduct AI Research Experiments? · On the Step Length Confounding in LLM Reasoning Data Selection 69.3 75 94%
17 University of Pennsylvania A Reproducible Evaluation of ANTs Similarity Metric Performance in Brain Image Registration · A Reproducible Evaluation of ANTs Similarity Metric Performance in Brain Image Registration · DECIPHERING SCIENTIFIC COLLABORATION IN BIOMEDICAL LLM RESEARCH: DYNAMICS, INSTITUTIONAL PARTICIPATION, AND RESOURCE DISPARITIES 69.3 52 93%
18 University of California Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering · ConvRML: High-Quality Lensless Imaging with Random Multi-Focal Lenslets · Hidden in Plain Text: Measuring LLM Deception Quality Against Human Baselines Using Social Deduction Games 69.26 223 94%
19 Northwestern Polytechnical University Toward Real-world Infrared Image Super-Resolution: A Unified Autoregressive Framework and Benchmark Dataset · SenFlow: Inter-Sentence Flow Modeling for AI-Generated Text Detection in Hybrid Documents · OODEval: Evaluating Large Language Models on Object-Oriented Design 69.26 49 92%
20 Nanyang Technological University VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding · Two-Stage Reinforcement Learning for Sound and Adversarial Test Generation in Code LLMs · R4-CGQA: Retrieval-based Vision Language Models for Computer Graphics Image Quality Assessment 69.25 165 93%
21 Southeast University SpikACom: A Neuromorphic Computing Framework for Green Communications · OmniGAIA: Towards Native Omni-Modal AI Agents · Hybrid guided variational autoencoder for visual place recognition 69.22 67 94%
22 Beihang University ScenePilot-Bench: A Large-Scale Dataset and Benchmark for Evaluation of Vision-Language Models in Autonomous Driving · NTIRE 2025 Challenge on UGC Video Enhancement: Methods and Results · LLMs-Powered Accurate Extraction, Querying and Intelligent Management of Literature-derived 2D Materials Data 69.2 84 93%
23 Technical University of Munich Coverage-Guided Road Selection and Prioritization for Efficient Testing in Autonomous Driving Systems · LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training · The MICCAI Federated Tumor Segmentation (FeTS) Challenge 2024: Efficient and Robust Aggregation Methods for Federated Learning 69.18 86 94%
24 University of Electronic Science and Technology of China Decoding the Sequence Determinants of Locus-Specific DNA Methylation Across Human Tissues · Holtercare-Bench: A Multimodal Benchmark for Evaluating Long-Term Dynamic ECG Analysis · Spiking Variational Graph Representation Inference for Video Summarization 69.16 67 94%
25 City University of Hong Kong DALD-PCAC: Density-Adaptive Learning Descriptor for Point Cloud Lossless Attribute Compression · Approaching Low-Cost Cardiac Intelligence with Semi-Supervised Knowledge Distillation · Metamorphic Testing for Audio Content Moderation Software 69.14 70 92%
26 The Hong Kong University of Science and Technology Metamorphic Testing for Audio Content Moderation Software · HiMed: Incentivizing Hindi Reasoning in Medical LLMs · MegaStyle: Constructing Diverse and Scalable Style Dataset via Consistent Text-to-Image Style Mapping 69.1 162 94%
27 Harbin Institute of Technology High-Quality Proposal Encoding and Cascade Denoising for Imaginary Supervised Object Detection · Spiking Variational Graph Representation Inference for Video Summarization · UDPNet: Unleashing Depth-based Priors for Robust Image Dehazing 69.09 119 94%
28 The Hong Kong Polytechnic University NTIRE 2025 Challenge on UGC Video Enhancement: Methods and Results · RESTORATION ADAPTATION FOR SEMANTIC SEGMENTATION ON LOW QUALITY IMAGES · High-resolution Photo Enhancement in Real-time: A Laplacian Pyramid Network 69.07 66 92%
29 Carnegie Mellon University CONCUR: BENCHMARKING LLMS FOR CONCURRENT CODEGENERATION · HOTPOTQA: A Dataset for Diverse, Explainable Multi-hop Question Answering · OMTRA: A Multi-Task Generative Model for Structure-Based Drug Design 69.04 105 94%
30 Johns Hopkins University Transfer Learning from One Cancer to Another via Deep Learning Domain Adaptation · RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology · SUMMARY OF THE INAUGURAL MUSIC SOURCE RESTORATION CHALLENGE 69.02 73 94%
31 Imperial College SpikACom: A Neuromorphic Computing Framework for Green Communications · The MICCAI Federated Tumor Segmentation (FeTS) Challenge 2024: Efficient and Robust Aggregation Methods for Federated Learning · ContractScrub: A benchmark for final review of legal contracts 69.01 85 92%
32 Huazhong University of Science and Technology WUTDet: A 100K-Scale Ship Detection Dataset and Benchmarks with Dense Small Objects · High-resolution Photo Enhancement in Real-time: A Laplacian Pyramid Network · Relit-LiVE: Relight Video by Jointly Learning Environment Video 68.97 65 93%
33 Beijing Institute of Technology Data Science and Technology Towards AGI Part I: Tiered Data Management · Holtercare-Bench: A Multimodal Benchmark for Evaluating Long-Term Dynamic ECG Analysis · MM-DETR: An Efficient Multimodal Detection Transformer with Mamba-Driven Dual-Granularity Fusion and Frequency-Aware Modality Adapters 68.97 63 92%
34 The University of Hong Kong MatTools: Benchmarking Large Language Models for Materials Science Tools · LLM-based Multimodal Feedback Produces Equivalent Learning and Better Student Perceptions than Educator Feedback · ARTEMIS: Autoregressive End-to-End Trajectory Planning with Mixture of Experts for Autonomous Driving 68.95 80 92%
35 Renmin University of China OmniGAIA: Towards Native Omni-Modal AI Agents · Metamorphic Testing for Audio Content Moderation Software · Feature Slice Matching for Precise Bug Detection 68.95 44 92%
36 Dalian University of Technology Toward Real-world Infrared Image Super-Resolution: A Unified Autoregressive Framework and Benchmark Dataset · Revisiting Salient Object Detection from an Observer-Centric Perspective · DCEvo: Discriminative Cross-Dimensional Evolutionary Learning for Infrared and Visible Image Fusion 68.85 42 92%
37 Northeastern University NTIRE 2025 Challenge on UGC Video Enhancement: Methods and Results · Scaling up fine-grained intracranial vessel annotations in computed tomography angiography · From Noisy Historical Maps to Time-Series Oil Palm Mapping Without Annotation in Malaysia and Indonesia (2020–2024) 68.78 69 93%
38 University of Cambridge SurvSurf: a partially monotonic neural network for first-hitting time prediction of intermittently observed discrete and continuous sequential events · ALPaCA: Adapting Llama for Pathology Context Analysis to enable slide-level question answering · PMPBench: A Paired Multi-Modal Pan-Cancer Benchmark for Medical Image Synthesis 68.76 88 94%
39 Xi’an Jiaotong University SAMIC: A Lightweight Semantic-Aware Mamba for Efficient Perceptual Image Compression · NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration: Methods and Results · Pan-denoising: Guided Hyperspectral Image Denoising via Weighted Represent Coefficient Total Variation 68.75 45 92%
40 New York University Solaris: Building a Multiplayer Video World Model in Minecraft · EMBOMATRIX: A SCALABLE TRAINING-GROUND FOREMBODIED DECISION-MAKING · VTaMo: Video-Text Alignment Model for Sign Language Translation 68.74 85 92%
41 University of Illinois Urbana-Champaign CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation Models · T-REN: Learning Text-Aligned Region Tokens Improves Dense Vision-Language Alignment and Scalability · NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration: Methods and Results 68.73 65 92%
42 University of Washington Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction · LandmarkLens: Predicting and Presenting Efective Landmarks for Mixed-Reality Urban Exploration · CLAIMDB: A Fact Verification Benchmark over Large Structured Data 68.72 62 94%
43 Nanjing University of Science and Technology Decoupling Continual Semantic Segmentation · AIM 2025 challenge on Inverse Tone Mapping Report: Methods and Results · The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report 68.67 44 92%
44 Monash University High-Quality Proposal Encoding and Cascade Denoising for Imaginary Supervised Object Detection · PanopMamba: Vision State Space Modeling for Nuclei Panoptic Segmentation · Restormer: Efficient Transformer for High-Resolution Image Restoration 68.62 48 92%
45 University of Oxford Connecting the Dots: A Machine Learning Ready Dataset for Ionospheric Forecasting Models · Self-Induced Outcome Potential: Turn-Level Credit Assignment for Agents without Verifiers · FDIF: Formula-Driven Supervised Learning with Implicit Functions for 3D Medical Image Segmentation 68.6 84 92%
46 Cornell University Tracking and Understanding Object Transformations · The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · Encoder-Only Image Registration 68.52 74 92%
47 University of Toronto DRIVINGGEN: A COMPREHENSIVE BENCHMARK FOR GENERATIVE VIDEO WORLD MODELS IN AUTONOMOUS DRIVING · LoRAFusion: Efficient LoRA Fine-Tuning for LLMs · LLM Safety From Within: Detecting Harmful Content with Internal Representations 68.46 70 92%
48 Tongji University An interactive enhanced driving dataset for autonomous driving · ScenePilot-Bench: A Large-Scale Dataset and Benchmark for Evaluation of Vision-Language Models in Autonomous Driving · HiMed: Incentivizing Hindi Reasoning in Medical LLMs 68.35 61 93%
49 Beijing University of Posts and Telecommunications SIDME: Self-supervised Image Demoiréing via Masked EncoderDecoder Reconstruction · Do Maps Still Matter for Machines: Revisiting the Role of Choropleth Maps in Foundation Model Spatial Understanding · Boosting Robustness for All-Weather Self-Supervised Depth Estimation in Autonomous Driving 68.29 54 94%
50 Columbia University A Reproducible Evaluation of ANTs Similarity Metric Performance in Brain Image Registration · A Reproducible Evaluation of ANTs Similarity Metric Performance in Brain Image Registration · Visual Instruction Tuning 68.13 59 94%

排名用于发现值得进一步核查的论文,不替代同行评审。评分会随新增引用、复现、实验验证、新闻评价和落地证据持续更新。

社区反馈如何参与?

PaperMiner MCP已开放结构化论文反馈。可提交复现成功或失败、代码可运行或失效、Benchmark纠错、链接纠错和部署结果。

不直接改分反馈独立形成“社区信号”,与六维证据评分隔离。
必须带依据需要公开链接或实验结果哈希;无依据的点赞不进入系统。
自动交叉验证至少两个独立账号给出同类结论后,才标记为已有多人印证。
自动防刷个人密钥限频、去重;新账号和利益相关反馈自动降低权重。

MCP接口:submit_paper_feedbackget_paper_community_feedback

权限:有效PaperMiner账号获配的个人MCP密钥可提交;公共代理和共享只读密钥只能查询。每个账号每天最多提交5条新反馈。