✦FireflyAI news, ranked by what matters
✦ Night SkyListLatest
✦ TopLatest
Model ReleaseFundingRegulationResearchInfrastructureProductOpinion
7
InfrastructurearXiv cs.AI·8d agoPrimary

Edge Physical AI Deployment of Vision Transformers on Heterogeneous Edge GPU Targeting Autonomous Vehicles

A new scheduling method for heterogeneous edge GPUs aims to improve the efficiency and throughput of vision transformer models in autonomous vehicle applications.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

ActiveFly-Bench: Aligning Embodied Question Answering with Vision-Language-Action for Aerial Embodied Perception

ActiveFly-Bench is a new benchmark designed to evaluate aerial embodied perception by linking high-level task understanding, behavior planning, and low-level control for unmanned aerial vehicles.

Read at arxiv.org ↗
7
Model ReleasearXiv cs.AI·8d agoPrimary

Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models

Embodied-R1.5 is a new foundation model designed to unify embodied reasoning and physical intelligence using a large-scale dataset of 15 billion tokens.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Valid $\ne$ Necessary: Diagnosing Latent Inefficiency in Chain-of-Thought

The study identifies a critical inefficiency in Chain-of-Thought prompting where models generate redundant but logically valid reasoning steps that current evaluators fail to penalize.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Remembering Distinct Items, Not Tokens: A Learnable Dirichlet-Process Cache Between State-Space Models and Attention

This paper proposes a learnable Dirichlet-process cache that stores only novel inputs to bridge the gap between state-space models and attention mechanisms.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy

This paper outlines a deployment-focused pathway for transitioning medical AI agents from simple assistants to autonomous clinical systems.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

IdeaTrail: Full-Process Agent Trajectories for Scientific Ideation

IdeaTrail is a new dataset designed to capture the complete, multi-stage workflow of AI agents performing scientific ideation and research tasks.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Are LLMs Ready for Scientific Discovery? A Capability-Oriented Benchmark for AI Scientists

Researchers have introduced SDABench, a new benchmark designed to evaluate the scientific data analysis capabilities of large language models across six distinct areas.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks

The authors introduce SWE-MERA, a dynamic benchmark designed to mitigate data contamination and improve the evaluation of LLMs on software engineering tasks.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Knowledge Graphs Meet Graph Neural Networks: A Comprehensive Survey

This paper provides a comprehensive survey and a new two-level taxonomy of Graph Neural Network methodologies applied across knowledge graph technologies.

Read at arxiv.org ↗
7
InfrastructurearXiv cs.AI·8d agoPrimary

MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference

MemDecay is a new memory management technique that optimizes LLM agent inference by using region-aware KV cache eviction based on semantic structure.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Model Collapse: On Recursion, Noise, and Uncharted Machine Visions

This paper examines the phenomenon of model collapse from both engineering and creative perspectives, exploring how recursive training on AI-generated data affects model output.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Unlocking Parallelism in Autoregressive Language Models via Speculative Decoding with Progressive Tree Drafting

Researchers propose a new speculative decoding method that improves large language model inference speed by utilizing progressive tree drafting to better exploit parallel processing.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Mako: A Self-Evolving Agentic Operating System (SE-AOS) for Autonomous Web Exploitation

The Mako project introduces a self-evolving agentic operating system capable of autonomously synthesizing and testing new security exploits.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

The Ramanujan Challenge For AI

Researchers have introduced a new benchmark dataset of mathematical constant formulas to evaluate the advanced mathematical reasoning capabilities of AI systems.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

FARS: A Fully Automated Research System Deployed at Scale

The authors introduce FARS, a fully automated system that enables AI agents to conduct research, generate hypotheses, and write manuscripts at scale.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories

Agentic-DPO is a new training framework designed to optimize AI agent policies on expert trajectories by teaching them to avoid plausible mistakes rather than just imitating sequences.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

The Ebb and Flow of Multimodal Focus: Scheduling Visual Relay Windows for Grounded VLM Reasoning

This study analyzes the internal attention dynamics of vision-language models to explain why visual grounding degrades and proposes scheduling visual relay windows to stabilize reasoning.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

HCRMap: Pressure-Aware Hot-Expert Residency Mapping for 3.5D MoE Chiplet Inference

The paper introduces HCRMap, a mapping framework designed to mitigate compute and communication imbalances caused by uneven expert activation in Mixture-of-Experts models on multi-chiplet systems.

Read at arxiv.org ↗
7
ResearcharXiv cs.AI·8d agoPrimary

Robust, Scalable Detection of Text Containment in Large Web-Crawled Corpora

Researchers released FindMyText, an open-source Python package designed to detect near-verbatim text containment within large web-crawled datasets using document fingerprinting.

Read at arxiv.org ↗
7
ProductQbitAI 量子位·8d ago

高德发布通用世界模型工坊ABot-World Studio:5090单卡可生成小时级实时交互式视频与3D场景

AutoNavi has launched ABot-World Studio, a tool that enables the generation of interactive 3D scenes and videos using a single consumer-grade GPU.

Read at qbitai.com ↗
7
ProductAWS ML Blog·8d ago

Launching UI for generative AI inference recommendations in Amazon SageMaker AI

Amazon has introduced a new user interface in SageMaker AI Studio to help developers easily benchmark and deploy optimized generative AI models.

Read at aws.amazon.com ↗
7
InfrastructureData Center Dynamics·5d ago

Managing Complexity: A Scalable Solution for Global Data Centers

A new framework has been introduced to assist data center operators in managing regulatory compliance and operational complexity on a global scale.

Read at datacenterdynamics.com ↗
7
ProductLeiphone 雷锋网·4d ago

独家丨智己回应经销商经营异常:需要时间妥善处理,厂家会兜底

IM Motors addressed concerns regarding dealer management issues following the launch of its new SUV model.

Read at leiphone.com ↗
← NewerPage 104Older →
Firefly aggregates headlines and links to original sources. All content belongs to its respective publishers.