✦FireflyAI news, ranked by what matters
✦ Night SkyListLatest
✦ TopLatest
Model ReleaseFundingRegulationResearchInfrastructureProductOpinion
6
ResearcharXiv cs.AI·6d agoPrimary

The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models

The paper provides a theoretical framework linking Joint-Embedding Predictive Architectures to active inference principles through variational free energy.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Traffic-Aware Randomized Smoothing for LLM-Based Network Intrusion Detection

A new defense mechanism called Traffic-Aware Randomized Smoothing has been proposed to improve the robustness of LLM-based network intrusion detection systems against traffic manipulation.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

MedDiffuseMix: Preserving Diagnostic Evidence with Saliency-Aware Diffusion Medical Image Data Augmentation

MedDiffuseMix provides a saliency-guided diffusion framework to augment medical imaging data while preserving critical diagnostic features.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

A Self-Evolving Agent for Longitudinal Personal Health Management

HealthClaw is a proposed self-evolving agent architecture designed to provide longitudinal health management by maintaining private memory of user routines and medical history.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

When Bots Join the Team: Bot Adoption and the Institutional Fabric of Open-Source Software Projects

Researchers analyzed nearly 3,000 GitHub projects to understand how the integration of automated bots as active participants influences the organizational structure of open-source software teams.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Unleashing Multimodal Large Language Models for Training-free HOI Detection in the Wild

Researchers have proposed a training-free approach for detecting human-object interactions in the wild by leveraging the capabilities of multimodal large language models.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies

A new method for test-time learning introduces learnable adaptation policies to help language agents improve their performance through iterative interaction.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Generative Compilation: On-the-Fly Compiler Feedback as AI Generates Code

The authors present a method for integrating compiler feedback directly into the autoregressive decoding process to improve the quality of AI-generated code.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

How Far Can Root Cause Analysis Go on Real-World Telemetry Data?

A study evaluating root cause analysis in microservice failures reveals that current AI and classical methods struggle to effectively process large-scale, multimodal telemetry data.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

A Causal Model of Theory of Mind in Conflict for Artificial Intelligence

This paper proposes a causal framework to determine when artificial intelligence systems should engage in theory of mind processes during conflict scenarios.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Set-shifting Behavioral Test for Harnessed Agents

A new benchmark based on cognitive psychology tests has been developed to evaluate how AI agents adapt when the reliability of their tools changes during operation.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

When Audio Separation Hurts Zero-Shot ASR: Evaluating SAM-Audio with Whisper on Bengali and English Speech

The study evaluates the impact of audio separation preprocessing on zero-shot automatic speech recognition performance, finding that cleaner audio does not always improve transcription accuracy.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Networked Intelligence: Active Shared Context Graphs for Human-AI Team Science

This paper proposes a framework for networked intelligence that uses shared context graphs to facilitate collaboration among multiple AI agents in scientific research.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

MASPRM: Multi-Agent System Process Reward Model

The Multi-Agent System Process Reward Model provides a method to evaluate and optimize message sequences between agents during inference-time search.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

DIVE: Embedding Compression via Self-Limiting Gradient Updates

DIVE is a new dimensionality reduction technique for language model embeddings that uses self-limiting gradient updates to improve compression efficiency.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Federated Explainable Artificial Intelligence: Roles, Architectures, Evaluation, and Open Challenges

This review examines the integration of explainable AI techniques within federated learning architectures to improve transparency in privacy-preserving distributed model training.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Experience Memory Graph: One-Shot Error Correction for Agents

A new error-correction method called Experience Memory Graph aims to help LLM agents recover from failures in long-horizon tasks more efficiently than traditional reflection techniques.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

STOCKTAKE: Measuring the Gap Between Perception and Action in LLM Agents with a Fair Oracle

A new evaluation framework called STOCKTAKE aims to distinguish between perception errors and execution failures in LLM agents during long-term decision-making tasks.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate

A study demonstrates that the interaction protocols used in multi-agent debates significantly influence the moral reasoning and judgment outcomes of large language models.

Read at arxiv.org ↗
6
ResearcharXiv cs.AI·6d agoPrimary

Interventional Grounding Audits: Black-Box Premise-Dependency Tests for LLM Chain-of-Thought via Predicate Substitution

A new auditing method uses predicate substitution to test whether large language models genuinely rely on stated premises during chain-of-thought reasoning.

Read at arxiv.org ↗
6
ProductData Center Dynamics·9d ago

Infineon, LS Electric partner on DC power infrastructure for AI data centers

Infineon and LS Electric have signed a non-binding memorandum of understanding to jointly develop next-generation direct current power systems for AI data centers.

Read at datacenterdynamics.com ↗
6
Funding36Kr 36氪·3d ago

中信证券与光谷金控签署框架合作协议

CITIC Securities and Optics Valley Financial Holdings have entered a partnership to provide capital market services and investment support for emerging technology companies.

Read at 36kr.com ↗
6
OpinionThe Verge AI·2d ago

SpaceX in your index fund, explained

This article discusses the potential financial risks and market implications of including SpaceX in major index funds.

Read at theverge.com ↗
6
ProductAWS ML Blog·9d ago

Building an agentic AI solution at Bluesight with Amazon Bedrock

Healthcare compliance company Bluesight has utilized Amazon Bedrock AgentCore to develop and deploy Prism, a unified agentic AI assistant.

Read at aws.amazon.com ↗
← NewerPage 111Older →
Firefly aggregates headlines and links to original sources. All content belongs to its respective publishers.