AY

Alan L. Yuille

cs.CVcs.LGeess.IVcs.AIcs.CLArtificial IntelligenceAlgorithmsNeural Networks, ComputerPancreatic Neoplasmscs.RO

On Valency

published · living versions
W_87gzj669·v1 · currentpublished
DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs
with Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy
1 version

Preprints & journals

495 papers in the corpus · 1986–2026
ImageNet3D: Towards General-Purpose Object-Level 3D Understanding.42445449 · Ma, Wufei, Zhang, Guofeng, Liu, Qihao et al.2026 · 4 citationsAdvances in neural information processing systems. 2024;37:96127-96149
Scaling 3D Compositional Models for Robust Classification and Pose Estimation.42088092 · Yuan, Xiaoding, Zhang, Guofeng, Kaushik, Prakhar et al.2026 · 0 citationsProceedings. IEEE International Conference on Computer Vision. 2025;2025:6406-6415
Scaling Laws in Patchification: An Image Is Worth 50,176 Tokens And More.41909270 · Wang, Feng, Yu, Yaodong, Shao, Wei et al.2026 · 3 citationsProceedings of machine learning research. 2025;267:65278-65290
Mamba-Reg: Vision Mamba Also Needs Registers.41394344 · Wang, Feng, Wang, Jiahao, Ren, Sucheng et al.2026 · 19 citationsProceedings. IEEE Computer Society Conference on Computer Vision and Pattern Recognition. 2025;2025:14944-14953
Adventurer: Optimizing Vision Mamba Architecture Designs for Efficiency.41179969 · Wang, Feng, Yang, Timing, Yu, Yaodong et al.2026 · 1 citationProceedings. IEEE Computer Society Conference on Computer Vision and Pattern Recognition. 2025;2025:30157-30166
Report Supervision.42664661 · Bassi, Pedro R A S, Li, Wenxuan, Wasserthal, Jakob et al.2026 · 0 citationsMedical image analysis. 2026;114:104257
Report Supervision2608.27668v1 · Pedro R. A. S. Bassia, Wenxuan Li, Jakob Wasserthal et al.2026 · 0 citationsarXiv
Exploiting Structural Consistency of Chest Anatomy for Unsupervised Anomaly Detection in Radiography Images.42672008 · Xiang, Tiange, Zhang, Yixiao, Lu, Yongyi et al.2026 · 16 citationsIEEE transactions on pattern analysis and machine intelligence. 2024;46(9):6070-6081
World-in-World: World Models in a Closed-Loop World2510.18135v2 · Jiahan Zhang, Muqing Jiang, Nanru Dai et al.2025 · 0 citationsarXiv
Large-Scale Multi-Cancer Detection by Learning Segmentation from Reports.42523553 · Bassi, Pedro R A S, Zhou, Xinze, Li, Wenxuan et al.2026 · 0 citationsResearch square. 2026
Development and evaluation of a computer vision algorithm for quantification of children's microactivities.41087742 · Lupolt, Sara N, Zhang, Guofeng, Wang, Jiahao et al.2026 · 2 citationsJournal of exposure science & environmental epidemiology. 2026;36(4):725-733
Hyperplasia Functions as a Link between Obesity and Cancer.41874564 · Pénisson, Sophie, Weischer, Maren, Li, Lu et al.2026 · 0 citationsCancer research. 2026;86(11):2678-2687
Application of a computer vision algorithm to quantify the frequency and duration of children's microactivities in different play scenarios.40796652 · Lupolt, Sara N, Lyu, Qinfan, Zhang, Guofeng et al.2026 · 0 citationsJournal of exposure science & environmental epidemiology. 2026;36(4):713-724
ARIA: A Causal-Aware Framework for Rescuing LLM Reasoning in Trustworthy Materials Discovery2606.22375v1 · Yi Cao, Liaoyaqi Wang, Jieneng Chen et al.2026 · 0 citationsProceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining V.2 (KDD '26), August 09--13, 2026, Jeju Island, Republic of Korea
A comprehensive survey of AI agents in healthcare.42009269 · Xu, Gelei, Li, Xueyang, Chen, Yixiong et al.2026 · 15 citationsJournal of biomedical informatics. 2026;179:105045
Text-Driven Tumor Synthesis.41605173 · Li, Xinran, Shuai, Yi, Liu, Chen et al.2026 · 4 citationsIEEE transactions on medical imaging. 2026;45(6):2639-2648
RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology2605.10761v1 · Wenxuan Li, Pedro R. A. S. Bassi, Xinze Zhou et al.2026 · 0 citationsarXiv
XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language Models2510.15148v2 · Xingrui Wang, Jiang Liu, Chao Huang et al.2025 · 1 citationPublished as a conference paper at ICLR 2026
Scaling Laws in Patchification: An Image Is Worth 50,176 Tokens And More2502.03738v2 · Feng Wang, Yaodong Yu, Guoyizhe Wei et al.2025 · 3 citationsarXiv
EigenLoRAx: Recycling Adapters to Find Principal Subspaces for Resource-Efficient Adaptation and Inference2502.04700v5 · Prakhar Kaushik, Ankit Vaidya, Shravan Chaudhari et al.2025 · 1 citationProceedings of the Computer Vision and Pattern Recognition Conference, 2025, pages 649-659
CoPE: Clipped RoPE as A Scalable Free Lunch for Long Context LLMs2602.05258v1 · Haoran Li, Sucheng Ren, Alan Yuille et al.2026 · 0 citationsarXiv
Shared LoRA Subspaces for almost Strict Continual Learning2602.06043v1 · Prakhar Kaushik, Ankit Vaidya, Shravan Chaudhari et al.2026 · 0 citationsarXiv
VTok: A Unified Video Tokenizer with Decoupled Spatial-Temporal Latents2602.04202v1 · Feng Wang, Yichun Shi, Ceyuan Yang et al.2026 · 0 citationsarXiv
Early and Prediagnostic Detection of Pancreatic Cancer from Computed Tomography2601.22134v1 · Wenxuan Li, Pedro R. A. S. Bassi, Lizhou Wu et al.2026 · 0 citationsarXiv
CausalSpatial: A Benchmark for Object-Centric Causal Spatial Reasoning2601.13304v1 · Wenxin Ma, Chenlong Wang, Ruisheng Yuan et al.2026 · 0 citationsarXiv
ReVision: Refining Video Diffusion with Explicit 3D Motion Modeling2504.21855v2 · Qihao Liu, Ju He, Qihang Yu et al.2025 · 0 citationsarXiv
Mixture of Contexts for Long Video Generation2508.21058v3 · Shengqu Cai, Ceyuan Yang, Lvmin Zhang et al.2025 · 1 citationarXiv
Expectation-Maximization as the Engine of Scalable Medical Intelligence2501.03410v2 · Wenxuan Li, Pedro R. A. S. Bassi, Tianyu Lin et al.2025 · 0 citationsarXiv
The Universal Weight Subspace Hypothesis2512.05117v2 · Prakhar Kaushik, Shravan Chaudhari, Ankit Vaidya et al.2025 · 0 citationsarXiv
Perceptual Taxonomy: Evaluating and Guiding Hierarchical Scene Reasoning in Vision-Language Models2511.19526v1 · Jonathan Lee, Xingrui Wang, Jiawei Peng et al.2025 · 0 citationsarXiv
TriDiff-4D: Fast 4D Generation through Diffusion-based Triplane Re-posing2511.16662v1 · Eddie Pokming Sheung, Qihao Liu, Wufei Ma et al.2025 · 0 citationsarXiv
Scaling Tumor Segmentation: Best Lessons from Real and Synthetic Data2510.14831v2 · Qi Chen, Xinze Zhou, Chen Liu et al.2025 · 5 citationsarXiv
Are Pixel-Wise Metrics Reliable for Sparse-View Computed Tomography Reconstruction?2506.02093v2 · Tianyu Lin, Xinran Li, Chuntung Zhuang et al.2025 · 0 citationsarXiv
Spatial457: A Diagnostic Benchmark for 6D Spatial Reasoning of Large Multimodal Models2502.08636v4 · Xingrui Wang, Wufei Ma, Tiezheng Zhang et al.2025 · 6 citationsarXiv
TGT: Text-Grounded Trajectories for Locally Controlled Video Generation2510.15104v1 · Guofeng Zhang, Angtian Wang, Jacob Zhiyuan Fang et al.2025 · 0 citationsarXiv
KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation2504.09656v2 · Xingrui Wang, Jiang Liu, Ze Wang et al.2025 · 0 citationsarXiv
Play to Generalize: Learning to Reason Through Game Play2506.08011v4 · Yunfei Xie, Yinsong Ma, Shiyi Lan et al.2025 · 0 citationsarXiv
Autoregressive Video Generation beyond Next Frames Prediction2509.24081v1 · Sucheng Ren, Chen Chen, Zhenbang Wang et al.2025 · 0 citationsarXiv
3DSRBench: A Comprehensive 3D Spatial Reasoning Benchmark2412.07825v4 · Wufei Ma, Haoyu Chen, Guofeng Zhang et al.2024 · 14 citationsarXiv
Generative World Explorer2411.11844v3 · Taiming Lu, Tianmin Shu, Alan Yuille et al.2024 · 0 citationsarXiv
RadGPT: Constructing 3D Image-Text Tumor Datasets2501.04678v2 · Pedro R. A. S. Bassi, Mehmet Can Yavuz, Kang Wang et al.2025 · 12 citationsarXiv
Career total: 1098 works. 495 are in this corpus.Showing the 50 most recent.

Profile built from the corpus for this byline.

Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.