AI.info
Latest Research — Page 44
Browse Latest Research on AI.info.
- Large-Scale Terminal Agentic Trajectory Generation from Dockerized Environments
- Domain Expansion: A Latent Space Construction Framework for Multi-Task Learning
- Generating Sketches in a Hierarchical Auto-Regressive Process for Flexible Sketch Drawing Manipulation at Stroke-Level
- How Many Tokens Do 3D Point Cloud Transformer Architectures Really Need?
- Text images processing system using artificial intelligence models
- UniMate: One Unified Model to Animate Diverse Skeletons
- Benchmarking Document Parsers on Mathematical Formula Extraction from PDFs
- Roleplaying with Structure: Synthetic Therapist-Client Conversation Generation from Questionnaires
- Right Tool, Right Job: Native-Language Evaluation, Tokenizer Sensitivity, and Methodological Findings from a French-Only BabyLM
- Adapting Noise to Data: Generative Flows from 1D Processes
- Dropouts in Confidence: Moral Uncertainty in Human-LLM Alignment
- SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
- LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling
- BurstDeflicker: A Benchmark Dataset for Flicker Removal in Dynamic Scenes
- RCScore: Quantifying Response Consistency in Large Language Models
- RS-ORT: A Reduced-Space Branch-and-Bound Algorithm for Optimal Regression Trees
- SigmaDock: Untwisting Molecular Docking With Fragment-Based SE(3) Diffusion
- Space Explanations of Neural Network Classification
- Causal Pre-training Under the Fairness Lens: An Empirical Study of TabPFN
- STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency Prediction
- K-Merge: Online Continual Merging of Adapters for On-device Large Language Models
- The Quest for Generalizable Motion Generation: Data, Model, and Evaluation
- ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body
- Schrödinger's Cat: Probabilistic Representation and Prediction of Potential Scene Kinematics
- Beyond Fixed Psychological Personas: State Beats Trait, but Language Models are State-Blind
- Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalization
- Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain
- Mitigating Privacy-Utility Trade-off in Decentralized Federated Learning via $f$-Differential Privacy
- ShotFinder: Imagination-Driven Open-Domain Video Shot Retrieval via Web Search
- Uni-Animator: Towards Unified Visual Colorization
- Zero-Shot Novel Depth Synthesis Using 3D Foundation Models Scene Representations
- 3D-Consistent Multi-View Editing by Correspondence Guidance
- Smooth trajectory generation and hybrid B-splines-Quaternions based tool path interpolation for a 3T1R parallel kinematic milling robot
- UG-UMRE: Uncertainty-Guided Modality Augmentation and Distributional Calibration for Unified Multimodal Relation Extraction
- Surgical Repair of Collapsed Attention Heads in ALiBi Transformers
- Enhancing Mathematical Problem Solving in LLMs through Execution-Driven Reasoning Augmentation
- MCPAgentBench: A Real-world Task Benchmark for Evaluating LLM Agent MCP Tool Use
- Watermarking for Factuality: Guiding Vision-Language Models Toward Truth via Tri-layer Contrastive Decoding
- Scalable Single-Cell Gene Expression Generation with Latent Diffusion Models
- CausalTrace: A Neurosymbolic Causal Analysis Agent for Smart Manufacturing