AY
Alan L. Yuille
cs.CVcs.LGeess.IVcs.AIcs.CLArtificial IntelligenceAlgorithmsNeural Networks, ComputerPancreatic Neoplasmscs.RO
On Valency
published · living versionsW_87gzj669·v1 · currentpublished
DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs
with Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy
1 version
Preprints & journals
495 papers in the corpus · 1986–2026ImageNet3D: Towards General-Purpose Object-Level 3D Understanding.42445449 · Ma, Wufei, Zhang, Guofeng, Liu, Qihao et al.2026 · 4 citationsAdvances in neural information processing systems. 2024;37:96127-96149
Scaling 3D Compositional Models for Robust Classification and Pose Estimation.42088092 · Yuan, Xiaoding, Zhang, Guofeng, Kaushik, Prakhar et al.2026 · 0 citationsProceedings. IEEE International Conference on Computer Vision. 2025;2025:6406-6415
Scaling Laws in Patchification: An Image Is Worth 50,176 Tokens And More.41909270 · Wang, Feng, Yu, Yaodong, Shao, Wei et al.2026 · 3 citationsProceedings of machine learning research. 2025;267:65278-65290
Mamba-Reg: Vision Mamba Also Needs Registers.41394344 · Wang, Feng, Wang, Jiahao, Ren, Sucheng et al.2026 · 19 citationsProceedings. IEEE Computer Society Conference on Computer Vision and Pattern Recognition. 2025;2025:14944-14953
Adventurer: Optimizing Vision Mamba Architecture Designs for Efficiency.41179969 · Wang, Feng, Yang, Timing, Yu, Yaodong et al.2026 · 1 citationProceedings. IEEE Computer Society Conference on Computer Vision and Pattern Recognition. 2025;2025:30157-30166
Report Supervision.42664661 · Bassi, Pedro R A S, Li, Wenxuan, Wasserthal, Jakob et al.2026 · 0 citationsMedical image analysis. 2026;114:104257
Report Supervision2608.27668v1 · Pedro R. A. S. Bassia, Wenxuan Li, Jakob Wasserthal et al.2026 · 0 citationsarXiv
Exploiting Structural Consistency of Chest Anatomy for Unsupervised Anomaly Detection in Radiography Images.42672008 · Xiang, Tiange, Zhang, Yixiao, Lu, Yongyi et al.2026 · 16 citationsIEEE transactions on pattern analysis and machine intelligence. 2024;46(9):6070-6081
World-in-World: World Models in a Closed-Loop World2510.18135v2 · Jiahan Zhang, Muqing Jiang, Nanru Dai et al.2025 · 0 citationsarXiv
Large-Scale Multi-Cancer Detection by Learning Segmentation from Reports.42523553 · Bassi, Pedro R A S, Zhou, Xinze, Li, Wenxuan et al.2026 · 0 citationsResearch square. 2026
Development and evaluation of a computer vision algorithm for quantification of children's microactivities.41087742 · Lupolt, Sara N, Zhang, Guofeng, Wang, Jiahao et al.2026 · 2 citationsJournal of exposure science & environmental epidemiology. 2026;36(4):725-733
Hyperplasia Functions as a Link between Obesity and Cancer.41874564 · Pénisson, Sophie, Weischer, Maren, Li, Lu et al.2026 · 0 citationsCancer research. 2026;86(11):2678-2687
Application of a computer vision algorithm to quantify the frequency and duration of children's microactivities in different play scenarios.40796652 · Lupolt, Sara N, Lyu, Qinfan, Zhang, Guofeng et al.2026 · 0 citationsJournal of exposure science & environmental epidemiology. 2026;36(4):713-724
ARIA: A Causal-Aware Framework for Rescuing LLM Reasoning in Trustworthy Materials Discovery2606.22375v1 · Yi Cao, Liaoyaqi Wang, Jieneng Chen et al.2026 · 0 citationsProceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining V.2 (KDD '26), August 09--13, 2026, Jeju Island, Republic of Korea
A comprehensive survey of AI agents in healthcare.42009269 · Xu, Gelei, Li, Xueyang, Chen, Yixiong et al.2026 · 15 citationsJournal of biomedical informatics. 2026;179:105045
Text-Driven Tumor Synthesis.41605173 · Li, Xinran, Shuai, Yi, Liu, Chen et al.2026 · 4 citationsIEEE transactions on medical imaging. 2026;45(6):2639-2648
DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents2605.09679v1 · Yixiong Chen, Wenjie Xiao, Pedro R. A. S. Bassi et al.2026 · 0 citationsarXiv
RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology2605.10761v1 · Wenxuan Li, Pedro R. A. S. Bassi, Xinze Zhou et al.2026 · 0 citationsarXiv
Distilling Photon-Counting CT into Routine Chest CT through Clinically Validated Degradation Modeling2604.07329v1 · Junqi Liu, Xinze Zhou, Wenxuan Li et al.2026 · 0 citationsarXiv
XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language Models2510.15148v2 · Xingrui Wang, Jiang Liu, Chao Huang et al.2025 · 1 citationPublished as a conference paper at ICLR 2026
Scaling Laws in Patchification: An Image Is Worth 50,176 Tokens And More2502.03738v2 · Feng Wang, Yaodong Yu, Guoyizhe Wei et al.2025 · 3 citationsarXiv
EigenLoRAx: Recycling Adapters to Find Principal Subspaces for Resource-Efficient Adaptation and Inference2502.04700v5 · Prakhar Kaushik, Ankit Vaidya, Shravan Chaudhari et al.2025 · 1 citationProceedings of the Computer Vision and Pattern Recognition Conference, 2025, pages 649-659
CoPE: Clipped RoPE as A Scalable Free Lunch for Long Context LLMs2602.05258v1 · Haoran Li, Sucheng Ren, Alan Yuille et al.2026 · 0 citationsarXiv
Shared LoRA Subspaces for almost Strict Continual Learning2602.06043v1 · Prakhar Kaushik, Ankit Vaidya, Shravan Chaudhari et al.2026 · 0 citationsarXiv
VTok: A Unified Video Tokenizer with Decoupled Spatial-Temporal Latents2602.04202v1 · Feng Wang, Yichun Shi, Ceyuan Yang et al.2026 · 0 citationsarXiv
Early and Prediagnostic Detection of Pancreatic Cancer from Computed Tomography2601.22134v1 · Wenxuan Li, Pedro R. A. S. Bassi, Lizhou Wu et al.2026 · 0 citationsarXiv
X-LRM: X-ray Large Reconstruction Model for Extremely Sparse-View Computed Tomography Recovery in One Second2503.06382v2 · Guofeng Zhang, Ruyi Zha, Hao He et al.2025 · 2 citationsarXiv
Large-Scale Label Quality Assessment for Medical Segmentation via a Vision-Language Judge and Synthetic Data2601.14406v1 · Yixiong Chen, Zongwei Zhou, Wenxuan Li et al.2026 · 0 citationsarXiv
CausalSpatial: A Benchmark for Object-Centric Causal Spatial Reasoning2601.13304v1 · Wenxin Ma, Chenlong Wang, Ruisheng Yuan et al.2026 · 0 citationsarXiv
ReVision: Refining Video Diffusion with Explicit 3D Motion Modeling2504.21855v2 · Qihao Liu, Ju He, Qihang Yu et al.2025 · 0 citationsarXiv
OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions2506.23361v3 · Yuanhao Cai, He Zhang, Xi Chen et al.2025 · 2 citationsarXiv
Artificial intelligence and radiologists in pancreatic cancer detection using standard of care CT scans (PANORAMA): an international, paired, non-inferiority, confirmatory, observational study.41275871 · Alves, Natalia, Schuurmans, Megan, Rutkowski, Dawid et al.2025 · 23 citationsThe Lancet. Oncology. 2026;27(1):116-124
Mixture of Contexts for Long Video Generation2508.21058v3 · Shengqu Cai, Ceyuan Yang, Lvmin Zhang et al.2025 · 1 citationarXiv
Expectation-Maximization as the Engine of Scalable Medical Intelligence2501.03410v2 · Wenxuan Li, Pedro R. A. S. Bassi, Tianyu Lin et al.2025 · 0 citationsarXiv
The Universal Weight Subspace Hypothesis2512.05117v2 · Prakhar Kaushik, Shravan Chaudhari, Ankit Vaidya et al.2025 · 0 citationsarXiv
Perceptual Taxonomy: Evaluating and Guiding Hierarchical Scene Reasoning in Vision-Language Models2511.19526v1 · Jonathan Lee, Xingrui Wang, Jiawei Peng et al.2025 · 0 citationsarXiv
TriDiff-4D: Fast 4D Generation through Diffusion-based Triplane Re-posing2511.16662v1 · Eddie Pokming Sheung, Qihao Liu, Wufei Ma et al.2025 · 0 citationsarXiv
Scaling Tumor Segmentation: Best Lessons from Real and Synthetic Data2510.14831v2 · Qi Chen, Xinze Zhou, Chen Liu et al.2025 · 5 citationsarXiv
Cue-Invariant Geometric Structure of the Population Codes in Macaque V1 and V210.1101/2023.12.05.570110v3 · Massot, C., Zhang, X., Wang, Z. et al.2023 · 0 citationsbioRxiv
Are Pixel-Wise Metrics Reliable for Sparse-View Computed Tomography Reconstruction?2506.02093v2 · Tianyu Lin, Xinran Li, Chuntung Zhuang et al.2025 · 0 citationsarXiv
Spatial457: A Diagnostic Benchmark for 6D Spatial Reasoning of Large Multimodal Models2502.08636v4 · Xingrui Wang, Wufei Ma, Tiezheng Zhang et al.2025 · 6 citationsarXiv
TGT: Text-Grounded Trajectories for Locally Controlled Video Generation2510.15104v1 · Guofeng Zhang, Angtian Wang, Jacob Zhiyuan Fang et al.2025 · 0 citationsarXiv
KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation2504.09656v2 · Xingrui Wang, Jiang Liu, Ze Wang et al.2025 · 0 citationsarXiv
Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction2411.14384v5 · Yuanhao Cai, He Zhang, Kai Zhang et al.2024 · 7 citationsarXiv
Gaussian Scenes: Pose-Free Sparse-View Scene Reconstruction using Depth-Enhanced Diffusion Priors2411.15966v3 · Soumava Paul, Prakhar Kaushik, Alan Yuille2024 · 0 citationsarXiv
Play to Generalize: Learning to Reason Through Game Play2506.08011v4 · Yunfei Xie, Yinsong Ma, Shiyi Lan et al.2025 · 0 citationsarXiv
Autoregressive Video Generation beyond Next Frames Prediction2509.24081v1 · Sucheng Ren, Chen Chen, Zhenbang Wang et al.2025 · 0 citationsarXiv
3DSRBench: A Comprehensive 3D Spatial Reasoning Benchmark2412.07825v4 · Wufei Ma, Haoyu Chen, Guofeng Zhang et al.2024 · 14 citationsarXiv
Generative World Explorer2411.11844v3 · Taiming Lu, Tianmin Shu, Alan Yuille et al.2024 · 0 citationsarXiv
RadGPT: Constructing 3D Image-Text Tumor Datasets2501.04678v2 · Pedro R. A. S. Bassi, Mehmet Can Yavuz, Kang Wang et al.2025 · 12 citationsarXiv
Career total: 1098 works. 495 are in this corpus.Showing the 50 most recent.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.