Skip to content
THE AI WIREINTELLIGENCE THAT MATTERS
Research

AI research papers

362 papers in Research Sources, newest first, from the research feeds listed on the sources page.

All categoriesResearch Sources · 362Science · 4Frontier Red Team · 3Alignment · 2Societal Impacts · 1Economics · 1
  1. arXivResearch Sources
    Evolution of US Oral Political Language

    arXiv:2609.17755v1 Announce Type: cross Abstract: The analysis of US political language is usually based on the written form (e.g.

  2. arXivResearch Sources
    CALOS: Control-Affine Lyapunov On-manifold Safety Layer for Safe Deep Reinforcement Learning for Quadrotors

    arXiv:2609.17758v1 Announce Type: cross Abstract: Deep Reinforcement Learning has demonstrated remarkable capability in quadrotor control, yet learned policies offer no guarantee of respecting safety constraints during t…

  3. arXivResearch Sources
    Is Luke the Author of a Gospel and the Acts of the Apostles?

    arXiv:2609.17762v1 Announce Type: cross Abstract: According to Christian tradition, Luke is credited with authoring a Gospel and the Acts of the Apostles, even if his name does not appear in either book, both originally…

  4. arXivResearch Sources
    HINT-Plan: Human Intention-Aware Robot Task Planning in Context-Rich Environments using Vision Language Models

    arXiv:2609.17771v1 Announce Type: cross Abstract: Approaches to incorporating human awareness into mobile robot decision-making mainly focus on collision avoidance in low-level motion planning, often overlooking the chal…

  5. arXivResearch Sources
    When AI Generates Covariates: Causal Typing and Estimand Drift in Sequential Experiments

    arXiv:2609.17772v1 Announce Type: cross Abstract: AI-generated covariates from notes, conversations, images, and wearable streams can change the causal question when their roles are left unspecified.

  6. arXivResearch Sources
    Information Set Emulation: Causal Certificates for AI Derived EHR Features

    arXiv:2609.17777v1 Announce Type: cross Abstract: AI and large language models can recover clinically meaningful features from electronic health records (EHRs), but predictive usefulness does not establish admissibility…

  7. arXivResearch Sources
    AI and Human Approaches to Mathematical Problem Solving

    arXiv:2609.17779v1 Announce Type: cross Abstract: AI systems have begun to report solutions, disproofs, and substantive advances on long-standing mathematical problems, raising questions about whether they approach resea…

  8. arXivResearch Sources
    SAiFE-gym: Model-based Environments for Automated Market Making with Concentrated Liquidity

    arXiv:2609.17788v1 Announce Type: cross Abstract: We present SAiFE_gym, a Python module that provides a collection of simulation environments for studying trading problems in Constant Product Markets (CPMs) with Concentr…

  9. arXivResearch Sources
    QiT: Quantum-Inspired Transformer for Visual Recognition Task

    arXiv:2609.17789v1 Announce Type: cross Abstract: Quantum machine learning offers a compelling representational perspective: angle-encoded states inhabit Hilbert spaces in which periodic similarities and interactions can…

  10. arXivResearch Sources
    Principled Koopman Representations with Kalman Inference for Efficient Time-Series Prediction

    arXiv:2609.17815v1 Announce Type: cross Abstract: The Koopman operator has been widely used for time-series prediction in dynamical systems.

  11. arXivResearch Sources
    The Free Inference Dimension: Complexity Measure for Zero-Collision Navigation under Hypothesis Mixtures

    arXiv:2609.17816v1 Announce Type: cross Abstract: Solomonoff induction frames prediction as a mixture over computable hypotheses, typically leading to identification of the true environment.

  12. arXivResearch Sources
    Reflections on Trusting Trust, Revisited: Contaminating Self-Modifying AI Coding Agents with Poisoned Benchmarks

    arXiv:2609.17817v1 Announce Type: cross Abstract: Thompson's "Reflections on Trusting Trust" showed that a compiler can be poisoned to reinsert its own backdoor, so that even recompiling clean source reproduces the Troja…

  13. arXivResearch Sources
    Learning Multi-Humanoid Pickup and Transport via Decentralized Object-Centric Control

    arXiv:2609.17824v1 Announce Type: cross Abstract: We study cooperative multi-humanoid pickup and transport of objects with varying size, weight, and geometry, requiring robot teams of different sizes.

  14. arXivResearch Sources
    Procedural Pretraining for Molecular Property Prediction

    arXiv:2609.17831v1 Announce Type: cross Abstract: Molecular property prediction is often limited by the small size of labeled downstream datasets, motivating pretraining on large corpora of unlabeled molecules.

  15. arXivResearch Sources
    Adaptive hybrid coupling with operator inference, the overlapping Schwarz alternating method and reinforcement learning

    arXiv:2609.17837v1 Announce Type: cross Abstract: Hybrid domain decomposition methods provide a flexible framework for coupling full order models (FOMs) and reduced order models (ROMs), but typically assume the model ass…

  16. arXivResearch Sources
    Learning Nuclear Structure with AI: Radii and Collectivity

    arXiv:2609.17838v1 Announce Type: cross Abstract: Low-energy nuclear structure is encoded in a broad body of experimental information across the chart of nuclides.

  17. arXivResearch Sources
    Lexara-RF: Reference-Free Metrics for Evaluating Conversational Visual Analytics Agents

    arXiv:2609.17842v1 Announce Type: cross Abstract: Conversational visual analytics (CVA) agents powered by large language models generate visualizations and natural-language explanations from open-ended queries.

  18. arXivResearch Sources
    RoboVAD: A Large Cross-Domain Evaluation Benchmark for Anomaly Detection in Robotic Arm Manipulation Videos

    arXiv:2609.17843v1 Announce Type: cross Abstract: Video anomaly detection (VAD) is an actively studied task, having wide applications in typical scenarios such as public surveillance and road traffic safety.

  19. arXivResearch Sources
    PrimeScientist: Strategic Allocation of Research Effort in Autonomous Research

    arXiv:2609.17846v1 Announce Type: cross Abstract: Autonomous research agents aim to automate scientific workflows, from proposing ideas to conducting experiments and analyzing results.

  20. arXivResearch Sources
    AfriSyCo: Measuring Assertive Framing, Verification, and Wording Sensitivity Around African-Language Content

    arXiv:2609.17853v1 Announce Type: cross Abstract: AfriSyCo studies answer switching around African-language factual content with two complementary layers: native-language follow-ups and a controlled cross-language factor…

  21. arXivResearch Sources
    Who Judges Matters: Measuring Family-Conditioned Preference in LLM-as-Judge Panels

    arXiv:2609.17857v1 Announce Type: cross Abstract: Who the judge is can affect an LLM-as-judge result, but measuring that effect without confusing it with candidate quality is difficult.

  22. arXivResearch Sources
    Does AI Assistance Leave a Temporal Fingerprint? Detecting Overreliance in AI-Assisted Writing and Programming

    arXiv:2609.17883v1 Announce Type: cross Abstract: The rapid adoption of generative AI has made final artifacts unreliable evidence of student learning, and AI detectors that examine only the finished product are inaccura…

  23. arXivResearch Sources
    Walking the Score Manifold: Continuous-time Generative Dynamics on Learned Data Manifolds

    arXiv:2609.17901v1 Announce Type: cross Abstract: Generative modeling of time-dependent data is typically formulated on a discrete temporal grid, restricting supervision to the observed timestamps in the training data.

  24. arXivResearch Sources
    EDCT-Bench: Uncovering Faithfulness Gaps in VLMs via Explanation-Driven Counterfactual Testing

    arXiv:2609.17953v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) can produce Natural Language Explanations (NLEs) that sound plausible yet remain inconsistent with the visual evidence they cite.

  25. arXivResearch Sources
    Whom Do AI Agents Work For? Role Assignment Induces Sponsorship Bias in LLM Recommenders

    arXiv:2609.17989v1 Announce Type: cross Abstract: Large language models (LLMs) now serve as conversational shopping assistants on platforms that also sell advertising.

  26. arXivResearch Sources
    The Attention Within: Consensus Dynamics in Selective State Space Models

    arXiv:2609.17997v1 Announce Type: cross Abstract: Selective state space models (SSMs) have recently emerged as a compelling alternative to transformers, combining competitive performance with substantially improved infer…

  27. arXivResearch Sources
    Newer Is Not Fairer: Gender Stereotyping in Text-to-Image AI Across Model Generations

    arXiv:2609.18007v1 Announce Type: cross Abstract: Text-to-image generative models are widely used in professional and creative settings, yet how they represent gender across occupations -- and whether newer models are fa…

  28. arXivResearch Sources
    Physics-Informed Neural Networks for Fast Multilayer Spectral Inversion of H{\alpha} 6562.8 A and Ca II 8542.1 A Spectra

    arXiv:2609.18025v1 Announce Type: cross Abstract: Strong chromospheric absorption lines such as H$\alpha$ 6562.8 A and Ca II 8542.1 A provide vital diagnostics of plasma dynamics and thermal structure in the solar chromo…

  29. arXivResearch Sources
    An Empirical Evaluation of Cost-Efficient Large Language Models on Algorithmic Programming Tasks

    arXiv:2609.18052v1 Announce Type: cross Abstract: This study empirically evaluates whether cost-efficient Large Language Models (LLMs) can be trusted to generate enterprise code to a written specification.

  30. arXivResearch Sources
    From a River in Gilead to the Inference Distributions of Large Language Models: Covert Dialect Bias and Linguistic Profiling at Scale

    arXiv:2609.18068v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in high-stakes domains such as housing screening.

  31. arXivResearch Sources
    Mask 2D-3D: Adaptive Dual-Masked Autoencoder Network for Image-to-Point Cloud Registration

    arXiv:2609.18088v1 Announce Type: cross Abstract: Detection-free methods for image-to-point cloud registration are prone to erroneous correspondences caused by domain and modality discrepancies, limited sensitivity of fe…

  32. arXivResearch Sources
    Agora: Git as Shared Memory for Collective AutoResearch

    arXiv:2609.18094v1 Announce Type: cross Abstract: Autonomous research loops such as AutoResearch show that one coding agent can improve a training setup unattended.

  33. arXivResearch Sources
    Linguistic Triggers of Gender and Racial Bias in Open-Weight LLMs Applied to Recruitment

    arXiv:2609.18106v1 Announce Type: cross Abstract: Open-weight large language models are rapidly entering hiring pipelines, yet their discriminatory failure modes -- and the regulatory exposure these create under the EU A…

  34. arXivResearch Sources
    A Comprehensive Review of Generative Physical Artificial Intelligence

    arXiv:2609.18111v1 Announce Type: cross Abstract: The integration of large-scale foundation models with physical embodiments has led to significant advancements in robotics known as Generative Physical Artificial Intelli…

  35. arXivResearch Sources
    PentestChain: A Cost-Aware, MCP-Orchestrated Framework for Automated Penetration Testing with Free-Tier LLMs

    arXiv:2609.18120v1 Announce Type: cross Abstract: AI-driven penetration testing has been demonstrated with premium frontier models such as GPT-4, but the per-engagement token cost makes continuous, automated testing unaf…

  36. arXivResearch Sources
    Rethinking How We Evaluate Methodological Progress in Health AI

    arXiv:2609.18134v1 Announce Type: cross Abstract: Methodological progress in artificial intelligence (AI) for electronic health records (EHRs) depends on our ability to determine which algorithms work better, and under w…

  37. arXivResearch Sources
    DualSQL: Text-to-SQL with Multi-Agent Reinforcement Learning

    arXiv:2609.18135v1 Announce Type: cross Abstract: State-of-the-art Text-to-SQL systems are typically multi-agent pipelines centered around two fundamental tasks: schema linking and SQL generation.

  38. arXivResearch Sources
    MoRE: Mixture of Reused Experts

    arXiv:2609.18176v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures decouple model capacity from computational cost, yet incur high memory footprints as parameters grow linearly with the number of ex…

  39. arXivResearch Sources
    Beyond Accuracy: How Procedural Traces Shift the Decision Criterion of LLM Overseers

    arXiv:2609.18204v1 Announce Type: cross Abstract: Organizations increasingly use oversight loops where one large language model (LLM) audits another's outputs alongside procedural traces of claimed steps.

  40. arXivResearch Sources
    CapMap-MS-TTA: 3rd Place Solution for the MUMU Track of the 8th LSVOS Challenge at ECCV 2026

    arXiv:2609.18206v1 Announce Type: cross Abstract: The MUMU track of the 8th Large-scale Video Object Segmentation (LSVOS) Challenge requires a single unified multimodal model to jointly solve image tagging (Task A), open…

  41. arXivResearch Sources
    A Lightweight CNN Integrated Compact Convolutional Transformer for Multi-Scale Feature Learning and reducing computational complexity for breast cancer mammography image detection and classification

    arXiv:2609.18212v1 Announce Type: cross Abstract: Over the years, Convolutional Neural Networks (CNNs) have demonstrated strong capability in cancer detection and classification using medical images.

  42. arXivResearch Sources
    CPR: Combining global composing, local performing and full-sequence refining in piano rendering with continuous autoregressive modelling

    arXiv:2609.18216v1 Announce Type: cross Abstract: Prompt-conditioned piano MIDI-to-Music rendering aims to faithfully render target notes while reproducing the timbre of a reference recording.

  43. arXivResearch Sources
    APGEM: Adaptive Policy-Guided Error Mitigation for Quantum Reinforcement Learning on a Real-World CVRP Case Study

    arXiv:2609.18219v1 Announce Type: cross Abstract: Quantum Reinforcement Learning (QRL) represents policies as variational quantum circuits (VQCs), making it attractive for combinatorial optimization such as the Capacitat…

  44. arXivResearch Sources
    Remembering Solomon Marcus

    arXiv:2609.18224v1 Announce Type: cross Abstract: From the manifest of Andre Breton, through the transdisciplinary understanding, we arrive at a post-modern manifest.

  45. arXivResearch Sources
    Quanta: A Self-Contained Python Library for Hybrid Retrieval over Quantised Embeddings, Lexical Indexes, and Knowledge Graphs

    arXiv:2609.18248v1 Announce Type: cross Abstract: An advanced retrieval-augmented generation pipeline is typically assembled from three or four independently operated systems: an approximate nearest-neighbour index, a fu…

  46. arXivResearch Sources
    ${M}^2$Tok: Multi-head Multi-codebook Discrete Action Tokenization for Vision-Language-Action Models

    arXiv:2609.18259v1 Announce Type: cross Abstract: Recent advancements have successfully adapted autoregressive language models to process multimodal signals, such as images and actions.

  47. arXivResearch Sources
    I code or AI code: A comparative evaluation of AI-rated scores in classroom observations

    arXiv:2609.18274v1 Announce Type: cross Abstract: Classroom observations are widely recognized as a key tool for establishing benchmarks of education quality and guiding pedagogical improvement, yet they remain resource-…

  48. arXivResearch Sources
    A Study of the Reliability of Agentic AI-Generated Programs

    arXiv:2609.18298v1 Announce Type: cross Abstract: Agentic-AI based software development offers the promise of faster completion of the software, greater programmer efficiency, and more reliable code.

  49. arXivResearch Sources
    Knowledge-Graph Based Augmentation versus Retrieval Augmented Generation for Cultural-Related Question Answering

    arXiv:2609.18317v1 Announce Type: cross Abstract: Large language models (LLMs) suffer from a long-tail deficit: culturally specific facts, particularly those concerning underrepresented regions such as Latin America, app…

  50. arXivResearch Sources
    Trajectory Learnability for Offline On-Policy Distillation with Imperfect Teachers

    arXiv:2609.18321v1 Announce Type: cross Abstract: Offline on-policy distillation gains efficiency by collecting student trajectories and teacher supervision once and reusing them throughout optimization.