✦FireflyAI news, ranked by what matters
✦ Night SkyListLatest
✦ TopLatest
Model ReleaseFundingRegulationResearchInfrastructureProductOpinion
7
ResearcharXiv cs.AI·7d agoPrimary

NL-PAC: Specification Ambiguity and Certified Minimax Risk Floors in LLM-Mediated Supervision

The NL-PAC framework is introduced to address and quantify the risks of specification ambiguity when using large language models for labeling and evaluation tasks.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·7d agoPrimary

Agora: Enhancing LLM Agent Reasoning Via Auction-Based Task Allocation

Agora is a proposed framework that uses an auction-based task allocation mechanism to optimize the coordination of diverse expert models and tools for LLM agents.

Read at arxiv.org ↗
7
ProductSimon Willison·3d ago

LLM cliché highlighter

A new utility tool has been released to identify and highlight common linguistic patterns and clichés frequently found in AI-generated text.

Read at simonwillison.net ↗
7
RegulationData Center Dynamics·3d ago

French competition watchdog to review SFR's €20.35bn carve up

French regulators have announced an 18-month review process regarding the proposed restructuring of the telecommunications company SFR.

Read at datacenterdynamics.com ↗
7
RegulationLeiphone 雷锋网·10d ago

OpenAI 权力洗牌:安全元老 Joshua Achiam 离职,白宫政策操盘手加入

OpenAI safety veteran Joshua Achiam has departed the company, while former White House policy strategist Dean Ball has joined as the head of strategic futures.

Read at leiphone.com ↗
7
ResearchMIT News AI·7d ago

How MIT students are helping to prevent cyberattacks

Students at the MIT Cybersecurity Clinic are assisting local governments and vulnerable groups in defending against cyber threats.

Read at news.mit.edu ↗
7
Model ReleaseOpenAI Blog·11d agoPrimary

GPT-5.6: Frontier intelligence that scales with your ambition

OpenAI has officially announced GPT-5.6, a new frontier model designed to offer higher intelligence per token and improved cost efficiency.

Read at openai.com ↗
7
ResearcharXiv cs.AI·6d agoPrimary

Boltzmann MapReduce: A Partition-Function Reduce for Forkable Sandboxes

The paper presents Boltzmann MapReduce, a framework that models distributed worker confidence outputs as Gibbs-Boltzmann measures to optimize partition-function reductions.

Read at arxiv.org ↗
7
ResearchQbitAI 量子位·3d ago

从仰望星空到落地创新:WAIC青年菁英会即将硬核开场,最新成果首发在即

The WAIC Youth Elite Forum showcased new AI research and development achievements from emerging industry talent.

Read at qbitai.com ↗
7
ResearcharXiv cs.AI·6d agoPrimary

Exploratory Analysis of Deep Learning Models for Forecasting Meteorological Parameters in the Agricultural Sector

This study evaluates and compares different recurrent and hybrid deep learning architectures for forecasting key meteorological parameters used in agricultural planning.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·6d agoPrimary

Prompting-MammAlps: Fine-Grained Text-to-Video Retrieval for Camera-Trap Data

Researchers introduced Prompting-MammAlps, a new benchmark designed to evaluate fine-grained text-to-video retrieval on ecological camera-trap datasets.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·6d agoPrimary

Constraint-Aware Hierarchical Search for Regulation-Driven Fine-Grained Classification

Researchers developed a constraint-aware hierarchical search method to improve fine-grained text classification under complex regulatory frameworks like customs tariffs and export controls.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·6d agoPrimary

BatteryLake: Agentic, Physics-Grounded Curation of Heterogeneous Battery Aging Data and Benchmarking

BatteryLake is a governed data lake that uses physics-grounded AI agents to automate the curation and standardization of diverse battery aging datasets.

Read at arxiv.org ↗
7
ProductThe Verge AI·8d ago

Apple’s failed self-driving car program left a legacy of powerful AI chips

Apple's cancelled self-driving car project ultimately drove the development of its Neural Engine, which now powers the company's on-device artificial intelligence processing.

Read at theverge.com ↗
7
Regulation36Kr 36氪·5d ago

英国将对十六七岁青少年实施社媒“宵禁”

The United Kingdom has announced a social media curfew for teenagers aged 16 and 17, restricting access between midnight and 6 a.m. and disabling features like autoplay by default.

Read at 36kr.com ↗
7
Funding36Kr 36氪·7d ago

近一周融资资金加码多只光模块龙头股,撤离存储等赛道

Recent Chinese market data shows that margin trading funds have increasingly targeted leading optical module stocks while pulling out of several semiconductor storage companies.

Read at 36kr.com ↗
7
ResearcharXiv cs.AI·4d agoPrimary

CAS I: A Geometric Coding Theorem

This theoretical paper introduces a geometric coding theorem that extends classical information theory concepts to symmetry groups.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·4d agoPrimary

Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable

The paper introduces a handbook approach to improve the maintainability and readability of AI agent harnesses as they evolve.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·4d agoPrimary

Rethinking Multimodal Fusion for Time Series: Text Modalities Need Constrained Fusion

This study demonstrates that integrating text modalities into time series forecasting requires constrained fusion strategies to achieve consistent performance improvements.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·4d agoPrimary

Stable Attention Response for Reliable Precipitation Nowcasting

A new approach to precipitation nowcasting focuses on improving the stability of attention responses in deep learning architectures.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·4d agoPrimary

Probabilistic Extension of Neuro-Symbolic AGI Robots based on Belnap's Typed Intensional FOL

This paper proposes a probabilistic extension to neuro-symbolic AI systems to improve logical reasoning and interpretability in robotic agents.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·4d agoPrimary

Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinforcement Learning Approach

A reinforcement learning approach is proposed to optimize signal coverage and energy efficiency in reconfigurable intelligent surfaces for edge computing.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·4d agoPrimary

Explaining Reinforcement Learning Agents via Inductive Logic Programming

A new study proposes using inductive logic programming to improve the interpretability of reinforcement learning policies in safety-critical applications.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·4d agoPrimary

Post-Training Pruning for Diffusion Transformers

Researchers have proposed a new post-training pruning method specifically designed to reduce the computational requirements of Diffusion Transformers.

Read at arxiv.org ↗
← NewerPage 89Older →
Firefly aggregates headlines and links to original sources. All content belongs to its respective publishers.