PAPER EVIDENCE RANKING
基于PaperMiner证据链动态评分。缺失证据保持未知,不等同于零分;不同任务论文不宜只凭总分直接比较。
| 排名 | 学校 | 得分 | 论文数 | 覆盖率 |
|---|---|---|---|---|
| 1 | Tsinghua University ScenePilot-Bench: A Large-Scale Dataset and Benchmark for Evaluation of Vision-Language Models in Autonomous Driving · OmniGAIA: Towards Native Omni-Modal AI Agents · NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results | 74.75 | 321 | 100% |
| 2 | 上海交通大学 Towards Visual Query Localization in the 3D World · NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results · Workspace-Bench 1.0: Benchmarking AI Agents on Workspace Tasks with Large-Scale File Dependencies | 74.75 | 274 | 100% |
| 3 | Zhejiang University RealHD: A High-Quality Dataset for Robust Detection of State-of-the-Art AI-Generated Images · NeurIDA: Dynamic Modeling for Effective In-Database Analytics · OmniGAIA: Towards Native Omni-Modal AI Agents | 74.75 | 246 | 100% |
| 4 | Peking University Efficient Content-based Recommendation Model Training via Noise-aware Coreset Selection · Crystal structure prediction with nuclear quantum and finite-temperature effects via deep free energy learning · FilDeep: Learning Large Deformations of Elastic-Plastic Solids with Multi-Fidelity Data | 74.75 | 227 | 100% |
| 5 | University of California L3GS: Layered 3D Gaussian Splats for Efficient 3D Scene Delivery · ConvRML: High-Quality Lensless Imaging with Random Multi-Focal Lenslets · Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering | 74.75 | 223 | 100% |
| 6 | University of Science and Technology of China RIVER: A REAL-TIME INTERACTION BENCHMARK FOR VIDEO LLMS · NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results · The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report | 74.75 | 183 | 100% |
| 7 | Nanyang Technological University NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results · RecNextEval: A Reference Implementation for Temporal Next-Batch Recommendation Evaluation · R4-CGQA: Retrieval-based Vision Language Models for Computer Graphics Image Quality Assessment | 74.75 | 165 | 95% |
| 8 | The Hong Kong University of Science and Technology Enhancing AIGC Service Efficiency with Adaptive Multi-Edge Collaboration in A Distributed System · LaVieID: Local Autoregressive Diffusion Transformers for Identity-Preserving Video Creation · Relit-LiVE: Relight Video by Jointly Learning Environment Video | 74.75 | 162 | 96% |
| 9 | Fudan University ABC-Bench: Benchmarking Agentic Backend Coding in Real-World Development · RIVER: A REAL-TIME INTERACTION BENCHMARK FOR VIDEO LLMS · FCMBench-Video: A Benchmark for Document-Video Intelligence | 74.75 | 154 | 100% |
| 10 | National University of Singapore Factorized Learning for Temporally Grounded Video-Language Models · NeurIDA: Dynamic Modeling for Effective In-Database Analytics · NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results | 74.75 | 147 | 100% |
| 11 | The Chinese University of Hong Kong Approaching Low-Cost Cardiac Intelligence with Semi-Supervised Knowledge Distillation · ReImagine: Rethinking Controllable High-Quality Human Video Generation via Image-First Synthesis · Relit-LiVE: Relight Video by Jointly Learning Environment Video | 74.75 | 145 | 98% |
| 12 | Stanford University CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation Models · The limits of fair medical imaging AI in real-world generalization · What LLMs Think When You Don’t Tell Them What to Think About? | 74.75 | 142 | 98% |
| 13 | University of Chinese Academy of Sciences Towards Visual Query Localization in the 3D World · A Mutil-conditional Diffusion Transformer for Versatile Seismic Wave Generation · Relit-LiVE: Relight Video by Jointly Learning Environment Video | 74.75 | 139 | 94% |
| 14 | Massachusetts Institute of Technology Workspace-Bench 1.0: Benchmarking AI Agents on Workspace Tasks with Large-Scale File Dependencies · The limits of fair medical imaging AI in real-world generalization · Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering | 74.75 | 138 | 97% |
| 15 | Harbin Institute of Technology High-Quality Proposal Encoding and Cascade Denoising for Imaginary Supervised Object Detection · NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration: Methods and Results · Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering | 74.75 | 119 | 96% |
| 16 | Carnegie Mellon University CONCUR: BENCHMARKING LLMS FOR CONCURRENT CODEGENERATION · LLM-based Multimodal Feedback Produces Equivalent Learning and Better Student Perceptions than Educator Feedback · HOTPOTQA: A Dataset for Diverse, Explainable Multi-hop Question Answering | 74.75 | 105 | 95% |
| 17 | Nanjing University RIVER: A REAL-TIME INTERACTION BENCHMARK FOR VIDEO LLMS · The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · GAP-URGENET: A GENERATIVE-PREDICTIVE FUSION FRAMEWORK FOR UNIVERSAL SPEECH ENHANCEMENT | 74.75 | 89 | 95% |
| 18 | University of Cambridge PMPBench: A Paired Multi-Modal Pan-Cancer Benchmark for Medical Image Synthesis · SurvSurf: a partially monotonic neural network for first-hitting time prediction of intermittently observed discrete and continuous sequential events · Iterative AI-guided optimisation of selective triple-drug combinations for breast cancer | 74.75 | 88 | 94% |
| 19 | Technical University of Munich VariViT: A Vision Transformer for Variable Image Sizes · MT4G: A Tool for Reliable Auto-Discovery of NVIDIA and AMD GPU Compute and Memory Topologies · MedSEBA: Synthesizing Evidence-Based Answers Grounded in Evolving Medical Literature | 74.75 | 86 | 96% |
| 20 | New York University Solaris: Building a Multiplayer Video World Model in Minecraft · Interactive Data Harmonization with LLM Agents: Opportunities and Challenges · Generative Recursive Reasoning | 74.75 | 85 | 96% |
| 21 | Imperial College SpikACom: A Neuromorphic Computing Framework for Green Communications · The MICCAI Federated Tumor Segmentation (FeTS) Challenge 2024: Efficient and Robust Aggregation Methods for Federated Learning · GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models | 74.75 | 85 | 95% |
| 22 | Beihang University ScenePilot-Bench: A Large-Scale Dataset and Benchmark for Evaluation of Vision-Language Models in Autonomous Driving · NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results · LLMs-Powered Accurate Extraction, Querying and Intelligent Management of Literature-derived 2D Materials Data | 74.75 | 84 | 94% |
| 23 | University of Oxford Connecting the Dots: A Machine Learning Ready Dataset for Ionospheric Forecasting Models · GMOS: Grounding Moving Object Segmentation in 3D Space and Time · Residual Decoder Adapter: ID-Preserving Tokenizer Adaption for Autoregressive Text Rendering | 74.75 | 84 | 93% |
| 24 | University College Compositional Segmentation of Cardiac Images Leveraging Metadata · Zero-shot Monocular Metric Depth for Endoscopic Images · INTRODUCTION AND NUMERICAL VALIDATION OF AN OPEN-SOURCE MATLAB PACKAGE FOR QUANTITATIVE ULTRASOUND TOMOGRAPHY VIA RAY-BORN INVERSION | 74.75 | 80 | 96% |
| 25 | The University of Hong Kong MatTools: Benchmarking Large Language Models for Materials Science Tools · LLM-based Multimodal Feedback Produces Equivalent Learning and Better Student Perceptions than Educator Feedback · Echo-Infinity: Learning Evolving Memory for Real-Time Infinite Video Generation | 74.75 | 80 | 93% |
| 26 | Wuhan University Towards Visual Query Localization in the 3D World · Universally Unfiltered and Unseen: Input-Agnostic Multimodal Jailbreaks against Text-to-Image Model Safeguards · Continual Vision-Language Learning for Remote Sensing: Benchmarking and Analysis | 74.75 | 78 | 96% |
| 27 | Sun Yat-sen University NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results · DepthAnything and SAM for UIE: Exploring Large Model Information Contributes to Underwater Image Restoration · DepthAnything and SAM for UIE: Exploring Large Model Information Contributes to Underwater Image Restoration | 74.75 | 76 | 94% |
| 28 | University of Michigan L3GS: Layered 3D Gaussian Splats for Efficient 3D Scene Delivery · Bi3: A Biplatform, Bicultural, Biperson Dataset for Social Robot Navigation · PRIVATEEDIT: A Privacy-Preserving Pipeline for Face-Centric Generative Image Editing | 74.75 | 75 | 94% |
| 29 | Cornell University The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · Tracking and Understanding Object Transformations · Credit-assigned Policy Gradient for Early Stage Retrieval in Two-stage Ranking | 74.75 | 74 | 94% |
| 30 | Johns Hopkins University RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology · SUMMARY OF THE INAUGURAL MUSIC SOURCE RESTORATION CHALLENGE · Large Variations Seen in First Ultraviolet Spectroscopic M33 Dust Extinction Curves | 74.75 | 73 | 95% |
| 31 | University of Toronto LoRAFusion: Efficient LoRA Fine-Tuning for LLMs · Reasoning About Reasoning: Towards Informed and Reflective Use of LLM Reasoning in HCI · Velox: Learning Representations of 4D Geometry and Appearance | 74.75 | 70 | 94% |
| 32 | City University of Hong Kong Approaching Low-Cost Cardiac Intelligence with Semi-Supervised Knowledge Distillation · Metamorphic Testing for Audio Content Moderation Software · DALD-PCAC: Density-Adaptive Learning Descriptor for Point Cloud Lossless Attribute Compression | 74.75 | 70 | 92% |
| 33 | Northeastern University The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · From Noisy Historical Maps to Time-Series Oil Palm Mapping Without Annotation in Malaysia and Indonesia (2020–2024) · NTIRE 2025 Challenge on UGC Video Enhancement: Methods and Results | 74.75 | 69 | 93% |
| 34 | Southeast University OmniGAIA: Towards Native Omni-Modal AI Agents · V2X-QA: A Comprehensive Reasoning Dataset and Benchmark for Multimodal Large Language Models in Autonomous Driving Across Ego, Infrastructure, and Cooperative Views · Dependency-Guided Repository-Level C-to-Rust Translation with Reinforcement Alignment | 74.75 | 67 | 96% |
| 35 | University of Electronic Science and Technology of China DRNet: All-in-One Image Restoration via Prior-Guided Dynamic Reparameterization · The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · AIM 2025 challenge on Inverse Tone Mapping Report: Methods and Results | 74.75 | 67 | 95% |
| 36 | The Hong Kong Polytechnic University High-resolution Photo Enhancement in Real-time: A Laplacian Pyramid Network · The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · RESTORATION ADAPTATION FOR SEMANTIC SEGMENTATION ON LOW QUALITY IMAGES | 74.75 | 66 | 93% |
| 37 | Huazhong University of Science and Technology High-resolution Photo Enhancement in Real-time: A Laplacian Pyramid Network · Source-Free Domain Adaptation (SFDA) for Privacy-Preserving Seizure Subtype Classification · Relit-LiVE: Relight Video by Jointly Learning Environment Video | 74.75 | 65 | 94% |
| 38 | University of Illinois Urbana-Champaign CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation Models · NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration: Methods and Results · Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection | 74.75 | 65 | 92% |
| 39 | Beijing Institute of Technology Data Science and Technology Towards AGI Part I: Tiered Data Management · MM-DETR: An Efficient Multimodal Detection Transformer with Mamba-Driven Dual-Granularity Fusion and Frequency-Aware Modality Adapters · Boosting Text-Driven Video Segmentation via Geometry-Aware Distillation | 74.75 | 63 | 92% |
| 40 | University of Washington The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction · Predicting Poets’ Origins from Verse: A Computational Analysis of Regional Linguistic Fingerprints in the Complete Tang Poems | 74.75 | 62 | 94% |
| 41 | University of Pennsylvania DECIPHERING SCIENTIFIC COLLABORATION IN BIOMEDICAL LLM RESEARCH: DYNAMICS, INSTITUTIONAL PARTICIPATION, AND RESOURCE DISPARITIES · A Reproducible Evaluation of ANTs Similarity Metric Performance in Brain Image Registration · The Brain Resection Multimodal Image Registration (ReMIND2Reg) 2025 Challenge | 74.75 | 52 | 94% |
| 42 | Northwestern Polytechnical University A Generative Data Framework with Authentic Supervision for Underwater Image Restoration and Enhancement · Toward Real-world Infrared Image Super-Resolution: A Unified Autoregressive Framework and Benchmark Dataset · OODEval: Evaluating Large Language Models on Object-Oriented Design | 74.75 | 49 | 93% |
| 43 | Monash University Human Factors in Immersive Analytics · High-Quality Proposal Encoding and Cascade Denoising for Imaginary Supervised Object Detection · Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering | 74.75 | 48 | 92% |
| 44 | Xi’an Jiaotong University Universally Unfiltered and Unseen: Input-Agnostic Multimodal Jailbreaks against Text-to-Image Model Safeguards · NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results · NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration: Methods and Results | 74.75 | 45 | 92% |
| 45 | Renmin University of China OmniGAIA: Towards Native Omni-Modal AI Agents · Graph Retrieval-Augmented Generation: A Survey · Metamorphic Testing for Audio Content Moderation Software | 74.75 | 44 | 94% |
| 46 | Nanjing University of Science and Technology NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results · The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · AIM 2025 challenge on Inverse Tone Mapping Report: Methods and Results | 74.75 | 44 | 92% |
| 47 | Dalian University of Technology The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report · Toward Real-world Infrared Image Super-Resolution: A Unified Autoregressive Framework and Benchmark Dataset · Revisiting Salient Object Detection from an Observer-Centric Perspective | 74.75 | 42 | 92% |
| 48 | Tongji University ScenePilot-Bench: A Large-Scale Dataset and Benchmark for Evaluation of Vision-Language Models in Autonomous Driving · An interactive enhanced driving dataset for autonomous driving · MAP: End-to-End Autonomous Driving with Map-Assisted Planning | 73.4 | 61 | 93% |
| 49 | Columbia University Visual Instruction Tuning · A Reproducible Evaluation of ANTs Similarity Metric Performance in Brain Image Registration · A Reproducible Evaluation of ANTs Similarity Metric Performance in Brain Image Registration | 73.4 | 59 | 94% |
| 50 | Purdue University PRIVATEEDIT: A Privacy-Preserving Pipeline for Face-Centric Generative Image Editing · FathomGPT: A Natural Language Interface for Interactively Exploring Ocean Science Data · Let’s Ask Gauss: Improved One-Run Privacy Auditing | 73.4 | 57 | 93% |
排名用于发现值得进一步核查的论文,不替代同行评审。评分会随新增引用、复现、实验验证、新闻评价和落地证据持续更新。
PaperMiner MCP已开放结构化论文反馈。可提交复现成功或失败、代码可运行或失效、Benchmark纠错、链接纠错和部署结果。
MCP接口:submit_paper_feedback、get_paper_community_feedback
权限:有效PaperMiner账号获配的个人MCP密钥可提交;公共代理和共享只读密钥只能查询。每个账号每天最多提交5条新反馈。