✦FireflyAI news, ranked by what matters
✦ Night SkyListLatest
✦ TopLatest
Model ReleaseFundingRegulationResearchInfrastructureProductOpinion
4
ResearcharXiv cs.AI·8d agoPrimary

Discrete Diffusion Models: A Unified Framework from Tokenization to Generation

A new framework for discrete diffusion models explores how tokenization and vocabulary structure influence generative performance.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

IMMNet: Hybrid Fusion of Model-based and Data-driven Approaches for Maneuvering Target Tracking

IMMNet integrates traditional model-based tracking algorithms with neural components to improve the accuracy and interpretability of maneuvering target tracking in 3D environments.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

ExTernD: Expanded-Rank Ternary Decomposition Ternary LLM PTQ with Accuracy Approaching Any Quantization Level

The ExTernD method introduces a post-training factorization technique for large language models that utilizes expanded-rank ternary decomposition to improve quantization efficiency and accuracy.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

GHR-VLM: Making Zero-Shot Transit Video Analytics Realizable with Grounded Hybrid Reasoning

The GHR-VLM framework combines grounded reasoning with vision-language models to enable zero-shot video analytics for transit systems without requiring task-specific training data.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

From Language to Navigation Goals: A Vision-Language Approach for Semantic Navigation of Mobile Robots Using RGB-D Perception

A new framework enables mobile robots to interpret natural language commands for autonomous navigation using RGB-D perception.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

Anatomically Faithful but Temporally Blind: Auditing Attribution for Left-Ventricular Ejection-Fraction Estimation from Echocardiography

A study evaluating deep learning models for echocardiography analysis reveals that current attribution methods often fail to verify temporal faithfulness in medical imaging predictions.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

Traffic-Aware Randomized Smoothing for LLM-Based Network Intrusion Detection

A new defense mechanism called Traffic-Aware Randomized Smoothing has been proposed to improve the robustness of LLM-based network intrusion detection systems against traffic manipulation.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate

A study demonstrates that the interaction protocols used in multi-agent debates significantly influence the moral reasoning and judgment outcomes of large language models.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

STOCKTAKE: Measuring the Gap Between Perception and Action in LLM Agents with a Fair Oracle

A new evaluation framework called STOCKTAKE aims to distinguish between perception errors and execution failures in LLM agents during long-term decision-making tasks.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

Experience Memory Graph: One-Shot Error Correction for Agents

A new error-correction method called Experience Memory Graph aims to help LLM agents recover from failures in long-horizon tasks more efficiently than traditional reflection techniques.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

DIVE: Embedding Compression via Self-Limiting Gradient Updates

DIVE is a new dimensionality reduction technique for language model embeddings that uses self-limiting gradient updates to improve compression efficiency.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

Federated Explainable Artificial Intelligence: Roles, Architectures, Evaluation, and Open Challenges

This review examines the integration of explainable AI techniques within federated learning architectures to improve transparency in privacy-preserving distributed model training.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·8d agoPrimary

Interventional Grounding Audits: Black-Box Premise-Dependency Tests for LLM Chain-of-Thought via Predicate Substitution

A new auditing method uses predicate substitution to test whether large language models genuinely rely on stated premises during chain-of-thought reasoning.

Read at arxiv.org ↗
4
ProductData Center Dynamics·10d ago

Infineon, LS Electric partner on DC power infrastructure for AI data centers

Infineon and LS Electric have signed a non-binding memorandum of understanding to jointly develop next-generation direct current power systems for AI data centers.

Read at datacenterdynamics.com ↗
4
Funding36Kr 36氪·5d ago

中信证券与光谷金控签署框架合作协议

CITIC Securities and Optics Valley Financial Holdings have entered a partnership to provide capital market services and investment support for emerging technology companies.

Read at 36kr.com ↗
4
OpinionThe Verge AI·3d ago

SpaceX in your index fund, explained

This article discusses the potential financial risks and market implications of including SpaceX in major index funds.

Read at theverge.com ↗
4
ProductAWS ML Blog·10d ago

Building an agentic AI solution at Bluesight with Amazon Bedrock

Healthcare compliance company Bluesight has utilized Amazon Bedrock AgentCore to develop and deploy Prism, a unified agentic AI assistant.

Read at aws.amazon.com ↗
4
RegulationData Center Dynamics·9d ago

Developer withdraws plan to replace hotel with data center in El Segundo, California

Developer Eight Form has withdrawn its proposal to replace a hotel with a data center in El Segundo, California, following public opposition during a planning meeting.

Read at datacenterdynamics.com ↗
4
ResearcharXiv cs.AI·10d agoPrimary

Partial Contracts Suffice: Sound, LLM-Inferred Regression Verification

Researchers have developed a contract-based regression verification tool that uses large language models to infer partial contracts, ensuring sound software patch verification without requiring manual specifications.

Read at arxiv.org ↗
4
ResearcharXiv cs.AI·10d agoPrimary

SPARK: Susceptibility-Guided Profiling and Steering of Latent Reasoning States in Large Language Models

The SPARK framework profiles and steers the latent reasoning states of large language models to diagnose and correct reasoning failures before final outputs are generated.

Read at arxiv.org ↗
4
ProductLeiphone 雷锋网·8d ago

一汽解放J6E大客户定制版批量交付,挚途L2++智驾系统成“快递神车”技术杀手锏

FAW Jiefang has delivered customized trucks equipped with Zhitu Technology's L2++ autonomous driving system for logistics operations.

Read at leiphone.com ↗
4
Opinion36Kr 36氪·11d ago

高盛:AI或引爆美国通胀,存储暴涨是核心推手

Goldman Sachs reports that the AI boom could trigger a wave of inflation in the US, driven primarily by supply constraints and rising prices for memory chips and semiconductors.

Read at 36kr.com ↗
4
Product36Kr 36氪·9d ago

中国稀土集团等在甘肃成立新公司

China Rare Earth Group has partnered with Gansu Guangsheng Rare Earth to establish a new joint venture focused on rare earth smelting and new material research.

Read at 36kr.com ↗
4
RegulationPandaily·10d ago

Why APEC 2026 Spotlight Turns to Chengdu: A Western Chinese City Digital Economy Transformation

Chengdu will host the APEC Digital and AI Ministerial Meeting in July 2026, showcasing its digital economy infrastructure and computing capacity.

Read at pandaily.com ↗
← NewerPage 139Older →
Firefly aggregates headlines and links to original sources. All content belongs to its respective publishers.