✦FireflyAI news, ranked by what matters
✦ Night SkyListLatest
✦ TopLatest
Model ReleaseFundingRegulationResearchInfrastructureProductOpinion
3
ResearcharXiv cs.AI·8d agoPrimary

Parameter-efficient Prompt Tuning of Vision Foundation Model With Adaptive Focal Loss for Interpretable MCI Screening

Researchers developed a parameter-efficient vision model tuning method to improve the automated screening of mild cognitive impairment from drawing tests.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Digital Pantheon: Simulating and Auditing Coalition Formation with LLM Agents

A multi-agent simulation framework has been created to study political coalition formation while bypassing standard LLM neutrality biases.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

CatalogAgent: A Supervisor-mediated Self-Learning System Enabling Context Engineering for GenAI Models

CatalogAgent is a new supervisor-mediated system designed to improve the accuracy of product attribute extraction in e-commerce databases using generative AI.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Measuring How Students Rely on Generative AI in Academic Writing: Development and Multi-Source Validation of the Generative AI Reliance Types Scale (GenAI-RTS)

A new psychometric scale has been developed to measure how undergraduate students rely on generative AI tools for academic writing tasks.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Eta Given Delta: Defining LLM Tool Efficiency With Marginal Tool Utility

A new quantitative metric has been developed to measure the efficiency and utility of tool calls within LLM agent trajectories.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Towards a Unified Multidimensional Explainability Metric: Evaluating Trustworthiness in AI Models

This paper introduces a multidimensional framework designed to standardize the evaluation of explainability methods in machine learning models.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Self-Evolving Human-Centered Framework for Explainable Depression Symptom Annotation

A new framework aims to improve the quality and explainability of depression symptom annotations by aligning them with clinical diagnostic standards.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

CoEvoT: Co-Evolving Chain-of-Thought Prompting for Graph-LLM Reasoning

The proposed CoEvoT framework enhances graph-based reasoning in large language models by co-evolving chain-of-thought prompts to better handle distribution shifts.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

ReportMedSAM: Guiding Segmentation Through Radiology Reports

ReportMedSAM utilizes natural language radiology reports to guide medical image segmentation, improving scalability and adaptability to new anatomical structures.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Volition Elicitation: Operational Semantics for People and Their Machines

The study introduces a formal framework for managing multi-agent transactions between humans and their personal devices based on user intent.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Local Additive Feature Attribution: A Mathematical Taxonomy and Reporting Checklist

This survey provides a unified mathematical taxonomy and reporting checklist for local additive feature attribution methods used in explainable AI.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Does generative AI supersede supervised XMLC? A Benchmark Study on Automated Subject Indexing with German Scientific Literature

This study compares the performance of generative AI against specialized supervised classification methods for the automated indexing of scientific literature.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

VLT: A Vision-Language-Time Series Multimodal Foundation Model for Industrial Intelligence

The VLT foundation model integrates vision, language, and time-series data to improve industrial equipment monitoring and prognostic health management.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

InCarEmo: A Multimodal Dataset for In-Cabin Emotion Recognition and Driver State Monitoring

Researchers released a multimodal dataset combining visual and conversational data to improve emotion recognition and state monitoring for automotive safety systems.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

ConFlow: Constraints-Guided Learning with Flow Matching for Motion Generation

The paper presents a flow-matching technique designed to incorporate specific constraints into robot motion generation tasks.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Reachability-Aware Pretraining for Efficient Target-Oriented Path Exploration in Temporal Knowledge Graph Reasoning

This paper proposes a reachability-aware pretraining method to improve the efficiency of reinforcement learning-based path exploration in temporal knowledge graph reasoning.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents?

A new study highlights that current performance-optimization benchmarks for coding agents may be unreliable due to issues with runtime instability and inconsistent scoring metrics.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

CFM-Bench: A Unified Multi-Domain, Multi-Task Benchmark for Channel Foundation Models

A new unified benchmark has been introduced to standardize the evaluation of foundation models applied to wireless communication tasks.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

Beyond scalar losses: calibrating segmentation models via gradient vector field surgery

This paper introduces a gradient vector field surgery technique to address the over-confidence and miscalibration issues common in medical image segmentation models trained with standard loss functions.

Read at arxiv.org ↗
3
ResearcharXiv cs.AI·8d agoPrimary

LBA: Textual Hard-Label Adversarial Attack under Low Query Budgets

A new method for generating adversarial text examples has been proposed to improve efficiency under low query budget constraints.

Read at arxiv.org ↗
3
Regulation36Kr 36氪·15d ago

遭连日舆论抨击后,Meta暂停AI图像生成功能

Meta has suspended an AI feature that allowed users to generate images using public Instagram photos following widespread criticism regarding its opt-out mechanism.

Read at 36kr.com ↗
3
RegulationThe Verge AI·15d ago

Apple sues OpenAI for allegedly stealing hardware secrets

Apple has filed a lawsuit against OpenAI and Jony Ive's hardware startup, alleging a systematic theft of trade secrets by former Apple employees.

Read at theverge.com ↗
3
RegulationTechCrunch AI·15d ago

Apple sues OpenAI over alleged trade secret theft

Apple is taking legal action against OpenAI, claiming the artificial intelligence company's leadership directed the theft of proprietary trade secrets.

Read at techcrunch.com ↗
3
ProductQbitAI 量子位·14d ago

中国首个十万卡集群落成!全国产算力支撑“十万卡时代”

China has completed its first 100,000-card GPU supercomputing cluster powered entirely by domestic hardware.

Read at qbitai.com ↗
← NewerPage 169Older →
Firefly aggregates headlines and links to original sources. All content belongs to its respective publishers.