Coverage: 2026-03-22 → 2026-03-29
We keep an eye on new AI papers on arXiv, pick one or two that really matter each day, and share the key ideas — no hype, just clear explanations.
Here’s what caught our eye over the past few days, unpacked by our AI-obsessed trio: Alex, your plain-language tech journalist host; Marc, the hands-on power user who’s built this stuff into real workflows; and Jamie, the senior AI/ML engineer making sure the models, integrations, and infra actually hold up in production.
LLM Daily – HeartAgent: An Autonomous Agent System for Explainable Differential Diagnosis in
Excerpt: Heart diseases remain a leading cause of morbidity and mortality worldwide, necessitating accurate and trustworthy differential diagnosis. However, existing artificial intelligence-based diagnostic methods are often…
LLM Daily – Natural-Language Agent Harnesses
Excerpt: Agent performance increasingly depends on harness engineering, yet harness design is usually buried in controller code and runtime-specific conventions, making it hard to transfer, compare, and study as a scientific…
LLM Daily – Schema on the Inside: A Two-Phase Fine-Tuning Method for High-Efficiency Text-to
Excerpt: Applying large, proprietary API-based language models to text-to-SQL tasks poses a significant industry challenge: reliance on massive, schema-heavy prompts results in prohibitive per-token API costs and high latency,…
LLM Daily – Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms
Excerpt: LLM agents like Claude Code can not only write code but also be used for autonomous AI research and engineering rank2026posttrainbench, novikov2025alphaevolve. We show that an autoresearch-style pipeline…
LLM Daily – Empirical Comparison of Agent Communication Protocols for Task Orchestration
Excerpt: Context. Nowadays, artificial intelligence agent systems are transforming from single-tool interactions to complex multi-agent orchestrations. As a result, two competing communication protocols have emerged: a tool…
LLM Daily – ProGRank: Probe-Gradient Reranking to Defend Dense-Retriever RAG from Corpus Poi
Excerpt: Retrieval-Augmented Generation (RAG) improves the reliability of large language model applications by grounding generation in retrieved evidence, but it also introduces a new attack surface: corpus poisoning. In this…
LLM Daily – PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and R
Excerpt: Empowering large language models with long-term memory is crucial for building agents that adapt to users' evolving needs. However, prior evaluations typically interleave preference-related dialogues with irrelevant…
LLM Daily – Designing Agentic AI-Based Screening for Portfolio Investment
Excerpt: We introduce a new agentic artificial intelligence (AI) platform for portfolio management. Our architecture consists of three layers. First, two large language model (LLM) agents are assigned specialized tasks: one…
LLM Daily – SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Pl
Excerpt: Agentic multimodal large language models (MLLMs) (e.g., OpenAI o3 and Gemini Agentic Vision) achieve remarkable reasoning capabilities through iterative visual tool invocation. However, the cascaded perception,…
LLM Daily – CatRAG: Functor-Guided Structural Debiasing with Retrieval Augmentation for Fair
Excerpt: Large Language Models (LLMs) are deployed in high-stakes settings but can show demographic, gender, and geographic biases that undermine fairness and trust. Prior debiasing methods, including embedding-space…
LLM Daily – CoverageBench: Evaluating Information Coverage across Tasks and Domains
Excerpt: We wish to measure the information coverage of an ad hoc retrieval algorithm, that is, how much of the range of available relevant information is covered by the search results. Information coverage is a central aspect…
LLM Daily – An Agentic Approach to Generating XAI-Narratives
Excerpt: Explainable AI (XAI) research has experienced substantial growth in recent years. Existing XAI methods, however, have been criticized for being technical and expert-oriented, motivating the development of more…
LLM Daily – Memori: A Persistent Memory Layer for Efficient, Context-Aware LLM Agents
Excerpt: As large language models (LLMs) evolve into autonomous agents, persistent memory at the API layer is essential for enabling context-aware behavior across LLMs and multi-session interactions. Existing approaches force…







































































Laisser un commentaire