✦FireflyAI news, ranked by what matters
✦ Night SkyListLatest
✦ TopLatest
Model ReleaseFundingRegulationResearchInfrastructureProductOpinion
2
ResearcharXiv cs.AI·13d agoPrimary

SAGEAgent: A Self-Evolving Agent for Cost-Aware Modality Acquisition in Multimodal Survival Prediction

SAGEAgent is a self-evolving agent designed to optimize the acquisition of clinical diagnostic modalities for cost-aware multimodal survival prediction in cancer patients.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·13d agoPrimary

CogniConsole: Externalizing Inference-Time Control as a Formal Abstraction for Reliable LLM Interactions

The paper introduces CogniConsole, an architecture that externalizes inference-time control into a structured interface to improve the reliability of large language model interactions.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Price of Fairness in Bandits: A Tight Minimax Characterization

The authors provide a theoretical characterization of the trade-offs between fairness and regret in bandit algorithms used for decision-making.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

NodeImport: Imbalanced Node Classification with Node Importance Assessment

The NodeImport method introduces a new approach to address class imbalance in graph neural networks by assessing node importance during classification tasks.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study

Researchers developed a multi-expert system using iterative fine-tuning to improve optical character recognition for historical Manchu documents with limited training data.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

EMAGN: Efficient Multi-Attention Graph Network via Learned Clustering for Scalable Traffic Forecasting

The proposed EMAGN architecture improves the scalability of traffic forecasting by using learned clustering to linearize spatial attention mechanisms.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Operator-on-F complements value-equivalence: a planning-time diagnostic for latent world models

Researchers introduced a diagnostic tool called operator-on-F to evaluate latent world models in reinforcement learning by comparing model predictions against environment observations.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·13d agoPrimary

Reward Transport: Property Control in Flow Matching via Noise-Space Alignment

The proposed Reward Transport framework utilizes optimal transport coupling during training to embed controllable structures directly into learned flow fields for molecular property control.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Grounded world models in biological organisms and future embodied AI

This paper explores the differences between current predictive AI training regimes and biological intelligence, arguing for the importance of grounded world models in future embodied AI development.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·13d agoPrimary

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors

Researchers analyzed evasion strategies against white-box monitors used for LLM safety and proposed new defensive mechanisms to improve auditing reliability.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Analyzing Curricular Pattern Complexity Using AI to Improve On-Time Graduation Rates

This study applies AI techniques to analyze academic curriculum patterns to help improve graduation rates in software engineering programs.

Read at arxiv.org ↗
2
Product36Kr 36氪·12d ago

联想在贵州成立新公司# 注册资本300万

Lenovo has established a new wholly-owned subsidiary in Guizhou with a registered capital of 3 million yuan to focus on AI hardware sales and software development.

Read at 36kr.com ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Disentangling Knowledge States with Ability and Proficiency Modeling for Knowledge Tracing

This paper proposes a knowledge tracing method that models student learning by disentangling ability and proficiency states from interaction sequences.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Explaining Reinforcement Learning Agents via Inductive Logic Programming

A new study proposes using inductive logic programming to improve the interpretability of reinforcement learning policies in safety-critical applications.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

CAS I: A Geometric Coding Theorem

This theoretical paper introduces a geometric coding theorem that extends classical information theory concepts to symmetry groups.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·13d agoPrimary

Are LLMs Ready to Assist Physicians? PhysAssistBench for Interactive Doctor-Patient-EHR Assistance

A new benchmark, PhysAssistBench, evaluates the ability of large language models to assist physicians by coordinating clinical knowledge, EHR interactions, and patient communication.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·13d agoPrimary

Video Generation Models are General-Purpose Vision Learners

This paper argues that large-scale text-to-video generation serves as an effective pre-training paradigm for developing general-purpose computer vision models.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

CoDiffGRN: Rethinking Gene Regulatory Network Inference via the BEELINE-KGC Benchmark and Co-evolutionary Discrete Diffusion

A new co-evolutionary diffusion model is proposed to improve the accuracy of gene regulatory network inference from single-cell data.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·13d agoPrimary

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation

Researchers introduced MedRealMM, a large-scale multimodal benchmark based on real-world Chinese online medical consultations to better evaluate LLMs in clinical practice.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·13d agoPrimary

LLM-Driven Evolutionary Generation of Multi-Objective Bayesian Optimization Algorithms

The LLaMEA framework has been extended to automatically generate complete multi-objective Bayesian optimization algorithms using large language models as evolutionary operators.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Explainable Artificial Intelligence for Anomaly Detection in Banking Transactions: An Internal Audit Perspective

A new explainable AI framework is proposed to assist banking internal audit teams in detecting and justifying anomalous transaction patterns.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinforcement Learning Approach

A reinforcement learning approach is proposed to optimize signal coverage and energy efficiency in reconfigurable intelligent surfaces for edge computing.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Partially Observed Structural Causal Models

The authors introduce Partially Observed Structural Causal Models to better account for upstream contexts in causal modeling frameworks.

Read at arxiv.org ↗
2
ResearcharXiv cs.AI·10d agoPrimary

Social Simulations: from Agent-Based Modeling to Digital Twins

This academic work explores the progression of social simulation techniques from traditional agent-based modeling to modern AI-driven digital twins.

Read at arxiv.org ↗
← NewerPage 186Older →
Firefly aggregates headlines and links to original sources. All content belongs to its respective publishers.