AI.info
Latest Research
Browse Latest Research on AI.info.
- Block Sparse Flash Attention
- Biases in the Blind Spot: Detecting What LLMs Fail to Mention
- Accelerating Vision Transformers with Adaptive Patch Sizes
- Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues
- Bootstrap Off-policy with World Model
- DeepCompress: A Dual Reward Strategy for Dynamically Exploring and Compressing Reasoning Chains
- VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection
- FirstAidQA: A Synthetic Dataset for First Aid and Emergency Response in Low-Connectivity Settings
- $μ$pscaling small models: Principled warm starts and hyperparameter transfer
- 3DPR: Single Image 3D Portrait Relight using Generative Priors
- Bilevel Layer-Positioning LoRA for Real Image Dehazing
- Stress Testing Factual Consistency Metrics for Long-Document Summarization
- Otter: Mitigating Background Distractions of Wide-Angle Few-Shot Action Recognition with Enhanced RWKV
- Feedback Control for Multi-Objective Graph Self-Supervision
- SVD-NO: Learning PDE Solution Operators with SVD Integral Kernels
- Finch: Benchmarking Finance & Accounting across Spreadsheet-Centric Enterprise Workflows
- iStructTab: Structured Feature Sequencing for Multimodal Learning of Image and Tabular Data
- Investigating the Presence and Development of Student Instructor Preferences in a Large-Scale CS1 Course
- NAUTILUS: A Large Multimodal Model for Underwater Scene Understanding
- Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer
- Multi-Aspect Mining and Anomaly Detection for Heterogeneous Tensor Streams
- NSR-Boost: A Neuro-Symbolic Residual Boosting Framework for Industrial Legacy Models
- BOTS: A Unified Framework for Bayesian Online Task Selection in LLM Reinforcement Finetuning
- Multi-Task GRPO: Reliable LLM Reasoning Across Tasks
- Political Advertising on Facebook During the 2022 Australian Federal Election: A Social Identity Perspective
- Creating ConLangs to Probe the Metalinguistic Grammatical Knowledge of LLMs
- Sketch2Colab: Sketch-Conditioned Multi-Human Animation via Controllable Flow Distillation
- ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding
- Disentangling Intrinsic Importance from Emergent Structure in Multi-Expert Orchestration
- Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs
- VecSet-Edit: Unleashing Pre-trained LRM for Mesh Editing from Single Image
- Omni-RRM: Advancing Omni Reward Modeling via Automatic Rubric-Grounded Preference Synthesis
- ConInstruct: Evaluating Large Language Models on Conflict Detection and Resolution in Instructions
- Automated Reproducibility Has a Problem Statement Problem
- How to Speculate about Uncertainty in Agentic Coding? A Draft-Model Gate Method
- AppellateGen: A Benchmark for Appellate Legal Judgment Generation
- Fast and Scalable Score-Based Kernel Calibration Tests
- LongT2IBench: A Benchmark for Evaluating Long Text-to-Image Generation with Graph-structured Annotations
- Decouple Searching from Training: Scaling Data Mixing via Model Merging for Large Language Model Pre-training
- Transformers Provably Learn Chain-of-Thought Reasoning with Length Generalization