AI.info
Latest Research — Page 11
Browse Latest Research on AI.info.
- Structure is Supervision: Multiview Masked Autoencoders for Radiology
- Weight-Space Mixture-of-Experts for Implicit Neural Representation Classification
- BLiSS 1.0: Evaluating Bilingual Learner Competence in Second Language Small Language Models
- Synchronization of Multiple Videos
- DiffFP: Learning Behaviors from Scratch via Diffusion-based Fictitious Play
- Text to Trust: Evaluating Fine-Tuning and LoRA Trade-offs in Language Models for Unfair Terms of Service Detection
- HATIR: Heat-Aware Diffusion for Turbulent Infrared Video Super-Resolution
- Post-Training Fairness Control: A Single-Train Framework for Dynamic Fairness in Recommendation
- SimWorld-Robotics: Synthesizing Photorealistic and Dynamic Urban Environments for Multimodal Robot Navigation and Collaboration
- UCO: A Multi-Turn Interactive Reinforcement Learning Method for Adaptive Teaching with Large Language Models
- ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs
- A Comprehensive Review of Bio-Inspired Approaches to Coordination, Communication, and System Architecture in Underwater Swarm Robotics
- Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
- Active Learning for Animal Re-Identification with Ambiguity-Aware Sampling
- Exposing Hidden Biases in Text-to-Image Models via Automated Prompt Search
- GROVER: Graph-guided Representation of Omics and Vision with Expert Regulation for Adaptive Spatial Multi-omics Fusion
- SCALE: Selective Resource Allocation for Overcoming Performance Bottlenecks in Mathematical Test-time Scaling
- EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning
- LAMP: Language-Assisted Motion Planning for Controllable Video Generation
- TALO: Pushing 3D Vision Foundation Models Towards Globally Consistent Online Reconstruction
- How Wide and How Deep? Mitigating Over-Squashing of GNNs via Channel Capacity Constrained Estimation
- Ming-UniVision: Joint Image Understanding and Generation with a Unified Continuous Tokenizer
- Are Language Models Efficient Reasoners? A Perspective from Logic Programming
- Rethinking Saliency Maps: A Cognitive Human Aligned Taxonomy and Evaluation Framework for Explanations
- IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
- Multiplicative Orthogonal Sequential Editing for Language Models
- Latent Action as Intention Enables Efficient Future Imagination for World Action Models
- ReIn: Conversational Error Recovery with Reasoning Inception
- TeamHOI: Learning a Unified Policy for Cooperative Human-Object Interactions with Any Team Size
- SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation
- UPLiFT: Efficient Pixel-Dense Feature Upsampling with Local Attenders
- How to Spend Your Oracle Budget: Practical Guidance for Protein Structure Prediction Models
- Grounding and Enhancing Informativeness and Utility in Dataset Distillation
- On the Collapse of Generative Paths: A Criterion and Correction for Diffusion Steering
- Echoing: Identity Failures when LLM Agents Talk to Each Other
- Forecasting in Offline Reinforcement Learning for Non-stationary Environments
- Docs2Synth: A Synthetic Data Trained Retriever Framework for Scanned Visually Rich Documents Understanding
- Small Foundation Models of Human Cognition and Behaviour
- DRIFT: Decompose, Retrieve, Illustrate, then Formalize Theorems
- Failure or Drift? Evaluating Monocular SLAM under Synthetic and Real-World Corruptions