✦FireflyAI news, ranked by what matters
✦ Night SkyListLatest
✦ TopLatest
Model ReleaseFundingRegulationResearchInfrastructureProductOpinion
4
ResearcharXiv cs.AI·9d agoPrimary

In-Context Reinforcement Learning under Non-Stationarity: A Survey

This survey paper reviews the current state of in-context reinforcement learning, focusing on how pretrained decision models handle non-stationary environments without parameter updates.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·9d agoPrimary

Explaining Data Mixing Scaling Laws

This research provides a theoretical framework to explain the mechanics behind data mixing scaling laws in multi-domain model training.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·9d agoPrimary

BAT-RM: A Boundary-Aware Transformer with Region-Aware Multi-Directional Mamba for Clinically Deployed Cervical Cancer Radiotherapy Auto-Contouring

A new hybrid transformer architecture is presented for automating cervical cancer radiotherapy contouring, demonstrating clinical deployment capabilities.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·9d agoPrimary

From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World

This study critiques current evaluation protocols for AI pentesting agents, arguing that existing benchmarks fail to accurately predict performance in real-world security environments.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·9d agoPrimary

Track, Rank, Crack: Epistemic Working Memory Scales Multi-Hop Reasoning in Language Agents

The SLEUTH framework improves multi-hop reasoning in language agents by explicitly managing epistemic working memory to prevent context dilution during long reasoning chains.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·9d agoPrimary

Less Experts, Faster Decoding: Cost-Aware Speculative Decoding for Mixture-of-Experts

A new speculative decoding strategy for Mixture-of-Experts models optimizes inference efficiency by accounting for the computational costs of specific expert activations.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·9d agoPrimary

GaitSpan: Growing Humanoid Locomotion from Walking to Running

The GaitSpan framework enables humanoid robots to transition between walking and running gaits without requiring separate training for each movement.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·3d agoPrimary

Learning-Driven Adaptive Audit Scheduling: A Sequential Decision Approach to Off-Chain Data Integrity

The authors present a deep reinforcement learning approach to optimize the scheduling of cryptographic audits for off-chain data storage.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·3d agoPrimary

Back to the museum: Investigation of the acceptance of Android Andrea with and without emotion simulation in a museum

A field study evaluates visitor acceptance of an autonomous multilingual android robot deployed in a public museum.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·3d agoPrimary

DMFNet: Dual-Backbone Multiscale Fusion Network for Urban Scene Classification

DMFNet is a proposed dual-backbone network architecture aimed at improving the classification of complex urban scenes in remote sensing imagery.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·3d agoPrimary

Position: Explanation Stability Is a Property of the Model Method Pair, Not the Model

A new position paper argues that explanation stability in machine learning models is tied to the specific model and interpretation method pair rather than the model alone.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·3d agoPrimary

Artificial Intelligence for Understanding and Managing Transportation Behavior in Sustainable Smart Cities

This paper explores the application of AI to analyze urban mobility data for the purpose of improving transportation management in smart cities.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·3d agoPrimary

OFD-Net: Teacher-Free Reliable Semi-supervised Medical Image Segmentation with Orthogonal Feature Disentanglement Net of Foreground-Background

A novel semi-supervised learning method called OFD-Net is introduced to improve medical image segmentation by utilizing orthogonal feature disentanglement without teacher models.

Read at arxiv.org ↗
4
OpinionQbitAI 量子位·6d ago

WAIC信息爆炸!大佬们都在说什么,笔记看这里

This article provides a summary of key insights and discussions from industry leaders and researchers at the World Artificial Intelligence Conference.

Read at qbitai.com ↗
4
ProductLeiphone 雷锋网·9d ago

百度搭子获评WAIC 2026“镇馆之宝”,智能体全家桶将集中亮相

Baidu's general-purpose AI agent, Baidu Dazi, has been recognized at the World Artificial Intelligence Conference 2026 for its productivity and task-automation capabilities.

Read at leiphone.com ↗
4
OpinionHugging Face Blog·9d ago

Model Routing Is Simple. Until It Isn’t.

This article explores the technical complexities and challenges involved in implementing model routing systems for AI applications.

Read at huggingface.co ↗
4
ProductData Center Dynamics·10d ago

Spain's Ferrovial to invest €1 billion in Madrid data center campus

Ferrovial is investing one billion euros to develop a new data center campus in Madrid with an initial 60MW capacity.

Read at datacenterdynamics.com ↗
4
FundingQbitAI 量子位·8d ago

工业母机进入“计算化时刻”:中国移动投资友机技术,押注工业AI下一代基础设施

China Mobile has invested in Youji Technology to advance the development of AI-integrated industrial infrastructure.

Read at qbitai.com ↗
4
Product36Kr 36氪·7d ago

宸境科技携系列核心产品亮相WAIC 2026

MirrorSense showcased new hardware and software tools at WAIC 2026 aimed at integrating spatial perception and robotics for industrial applications.

Read at 36kr.com ↗
4
ResearchLeiphone 雷锋网·9d ago

登顶 ICML Oral !专访上交大团队:这个 3D 自动标注 AI 太强了

Researchers have developed Holi-Spatial, an AI method for 3D automatic annotation that significantly improves detection accuracy without the need for manual labeling or expensive hardware.

Read at leiphone.com ↗
4
ResearcharXiv cs.AI·7d agoPrimary

Closed-Loop Knowledge Dynamics: An Operational Framework for Saturation and Escape

Researchers propose a theoretical framework to address performance saturation in closed-loop AI systems by identifying how external information can overcome internal feedback limitations.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·7d agoPrimary

Can LLMs Build a MaxSAT Solver from Papers? The CoreForge Experience

The CoreForge project demonstrates an iterative workflow where large language models are used to construct a MaxSAT solver directly from academic literature rather than existing codebases.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·7d agoPrimary

StructureClaw: Traceable LLM Agents and an Executable Benchmark for Structural Engineering Workflows

Researchers introduced a benchmark designed to evaluate the accuracy and traceability of LLM agents performing complex structural engineering workflows.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·7d agoPrimary

ANet Patu-1: The Value of Connection in the Agent Network

This paper models the value of connectivity within networks of AI agents to determine optimal collaboration protocols.

Read at arxiv.org ↗
← NewerPage 145Older →
Firefly aggregates headlines and links to original sources. All content belongs to its respective publishers.