✦FireflyAI news, ranked by what matters
✦ Night SkyListLatest
✦ TopLatest
Model ReleaseFundingRegulationResearchInfrastructureProductOpinion
3
ResearcharXiv cs.AI·11d agoPrimary

Quantum Circuit Vision: Cost-Aware Evaluation of Visual AI Agents for Quantum Code Generation

Quantum Circuit Vision is a new cost-aware benchmark designed to evaluate how well multimodal AI agents can interpret quantum circuit diagrams and generate executable code.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

ML in a Box: Analyzing Containerization Practices in Open Source ML Projects

This empirical study analyzes containerization practices in open-source machine learning projects to understand how iterative workflows affect build performance and container size.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Co4ICF: Co-evolving Physics-Informed Surrogate and RL-based Pulse Optimizer for Inertial Confinement Fusion

The Co4ICF framework couples a physics-informed surrogate model with a reinforcement learning optimizer to prevent out-of-distribution errors in inertial confinement fusion simulations.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Large Language Models in Misinformation Ecosystems: Misuse, Defense, and Vulnerability

A new study analyzes the systemic risks posed by large language models in misinformation ecosystems, proposing a framework to categorize vulnerabilities and defense strategies.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

DiffUE: Enhancing Utility-Unlearnability Trade-off of Unlearnable Examples via Diffusion Autoencoders

Researchers introduced DiffUE, a method using diffusion autoencoders to improve the effectiveness of unlearnable examples in protecting images from unauthorized AI training.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Distributed Denial of Science: How Indirect Data Poisoning of AI Systems Can Industrialize Scientific Fraud

This study explores the potential for malicious actors to automate scientific fraud by weaponizing AI systems to generate misleading research data.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Laguerre Geometry for Interpreting Large Language Models

This study proposes using Laguerre Geometry to model and analyze how concepts are structured and separated within large language models.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Subjective Risk Decomposition: A New View for Uncertainty Quantification

This paper introduces a theoretical framework for quantifying uncertainty in AI models by decomposing subjective risk based on loss functions.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Agentic Routing: The Harness-Native Data Flywheel

A new approach to agentic routing is proposed to optimize model selection within execution harnesses by leveraging the specialized strengths of different AI models.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Asymmetric Peak-Aware Loss for Peak-Critical Time Series Forecasting

A new loss function designed for time-series forecasting aims to improve the prediction of rare demand spikes by prioritizing extreme-value accuracy.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Technical Report on the CVPR 2026@AdvML Workshop Challenge

A technical report details a CVPR 2026 challenge focused on testing the robustness of autonomous driving vision-language agents against adversarial multimodal attacks.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias

A study into the mechanistic interpretability of LLM-as-a-judge models reveals that scoring biases can be identified and analyzed within the internal hidden states of the models.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Safety from Honesty in a Disinterested AI Predictor

This paper proposes a formal safety framework for an AI predictor designed to provide honest outputs by conditioning on epistemically contextualized data.

Read at arxiv.org ↗
3
Funding36Kr 36氪·10d ago

阿里巴巴获南向资金净买入约30.91亿港元

Alibaba and Tencent received significant net inflows of Southbound capital, totaling approximately 3.09 billion HKD and 1.81 billion HKD respectively.

Read at 36kr.com ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Private Seeds, Public LLMs: Realistic and Privacy-Preserving Synthetic Data Generation

Researchers developed a method called RPSG that utilizes private seeds and differential privacy to generate synthetic data while balancing privacy and utility.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Small edits, large models: How Wikipedia advocacy shapes LLM values

A study demonstrates that small-scale, targeted edits to Wikipedia articles can measurably influence the values and outputs of large language models trained on that data.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

Breaking the Quality--Intelligibility Trade-off in Streaming Target Speaker Extraction via Deep-Feature-Anchored Preference Optimization

A new optimization method called Deep-Feature-Anchored Preference Optimization resolves the trade-off between audio quality and speech intelligibility in streaming target speaker extraction models.

Read at arxiv.org ↗
3
InfrastructurearXiv cs.AI·11d agoPrimary

NaviCache: Test-Time Self-Calibration Caching for Video Generation

NaviCache introduces a calibration-free caching method to reduce the computational costs associated with generating video via diffusion models.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

SETA: Scaling Environments for Terminal Agents

Researchers have introduced SETA, a framework designed to scale the training of terminal-based language model agents by generating diverse and coherent task instructions.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

JEEVHITAA -- An HCAI Ecosystem to Support Collective Care

The JEEVHITAA platform introduces a mobile system designed to facilitate coordinated, multi-actor information sharing and workflows within healthcare settings.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

SWIFT: A Small-World Interaction Framework for Flow-Aware Trajectory Prediction in Autonomous Driving

Researchers developed SWIFT, a framework that incorporates structural priors from traffic networks to improve trajectory prediction and generalization in autonomous driving.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·11d agoPrimary

The Verifier is the Curriculum: Execution-Gated Self-Distillation for Cross-Family Game Generation

This paper introduces an execution-gated self-distillation method that uses a deterministic launch filter to improve the cross-family generalization of code generators in game development.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

ReportMedSAM: Guiding Segmentation Through Radiology Reports

ReportMedSAM utilizes natural language radiology reports to guide medical image segmentation, improving scalability and adaptability to new anatomical structures.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·5d agoPrimary

Candidate Attended Dialogue State Tracking Using BERT

This paper explores the application of BERT-based models to improve dialogue state tracking in task-oriented conversational AI systems.

Read at arxiv.org ↗
← NewerPage 166Older →
Firefly aggregates headlines and links to original sources. All content belongs to its respective publishers.