Researchers have made significant progress in developing stable aggregation methods for quantum federated learning, enabling clients to train quantum neural network models without sharing private data. A novel self-consistent midpoint aggregation method has been developed and validated through extensive evaluations and experiments on medical and financial datasets.
The method combines quantum-of-service-aware client weighting, circular parameter aggregation, and bounded midpoint-based update control, achieving improved stability, lower volatility, and competitive accuracy. The results have been published in a recent research article.
In another study, a conversational AI system called Conversation Coach has been proposed to help practice difficult workplace conversations. The system uses a voice-first AI approach to enable managers to rehearse conversations in a realistic spoken format. It has been compared with an end-to-end speech-to-speech model and a cascaded approach combining automatic speech recognition, a large language model, and text-to-speech synthesis.
The results show that the end-to-end approach achieves lower latency with native barge-in capability at a lower cost, while the cascaded approach offers superior reasoning essential for coaching quality. The system has been deployed in production and used by 40,000+ managers over six months.
A new method for compressing reasoning chains while preserving answer accuracy and logical coherence has been proposed. The Hierarchical Semantic Distillation Network (HSDN) framework combines semantic segmentation, dependency graph construction, dual encoder importance scoring, constrained segment selection, and local boundary rewriting. The results show that HSDN achieves 91.0% accuracy with 68.4% compression, outperforming strong compression baselines in overall score and reasoning coherence.
Researchers have developed a novel Multi-Agent Retrieval-Augmented Generation (RAG) system for spectrum intelligence, enabling autonomous agents to coordinate specialized sub-agents that retrieve and synthesize knowledge across policy proceedings, legal regulations, and license databases. The system has been evaluated on a question and answer (Q&A) dataset based on real-world license records and policy proceedings, and has achieved over 80% win rate against strong baselines.
A new framework for evaluating task-oriented dialogue agents has been proposed, which compiles a workflow specification and per-turn state diff into atomic, schema-grounded criteria and routes each through a cascade of symbolic and encoder/NLI verifiers. The framework has been evaluated on four slices spanning MultiWOZ, Schema-Guided Dialogue, and ABCD, and has shown strong performance compared with state-of-the-art baselines.
Researchers have developed a voice-enabled AI system called Conversation Coach to help practice difficult workplace conversations. The system uses a voice-first AI approach to enable managers to rehearse conversations in a realistic spoken format. It has been compared with an end-to-end speech-to-speech model and a cascaded approach combining automatic speech recognition, a large language model, and text-to-speech synthesis.
A new method for compressing reasoning chains while preserving answer accuracy and logical coherence has been proposed. The Hierarchical Semantic Distillation Network (HSDN) framework combines semantic segmentation, dependency graph construction, dual encoder importance scoring, constrained segment selection, and local boundary rewriting. The results show that HSDN achieves 91.0% accuracy with 68.4% compression, outperforming strong compression baselines in overall score and reasoning coherence.
Researchers have developed a novel Multi-Agent Retrieval-Augmented Generation (RAG) system for spectrum intelligence, enabling autonomous agents to coordinate specialized sub-agents that retrieve and synthesize knowledge across policy proceedings, legal regulations, and license databases. The system has been evaluated on a question and answer (Q&A) dataset based on real-world license records and policy proceedings, and has achieved over 80% win rate against strong baselines.
Key Takeaways
- Researchers have developed a stable aggregation method for quantum federated learning, enabling clients to train quantum neural network models without sharing private data.
- A novel self-consistent midpoint aggregation method has been developed and validated through extensive evaluations and experiments on medical and financial datasets.
- The method combines quantum-of-service-aware client weighting, circular parameter aggregation, and bounded midpoint-based update control, achieving improved stability, lower volatility, and competitive accuracy.
- A conversational AI system called Conversation Coach has been proposed to help practice difficult workplace conversations.
- The system uses a voice-first AI approach to enable managers to rehearse conversations in a realistic spoken format.
- A new method for compressing reasoning chains while preserving answer accuracy and logical coherence has been proposed.
- The Hierarchical Semantic Distillation Network (HSDN) framework combines semantic segmentation, dependency graph construction, dual encoder importance scoring, constrained segment selection, and local boundary rewriting.
- The results show that HSDN achieves 91.0% accuracy with 68.4% compression, outperforming strong compression baselines in overall score and reasoning coherence.
- Researchers have developed a novel Multi-Agent Retrieval-Augmented Generation (RAG) system for spectrum intelligence.
- The system enables autonomous agents to coordinate specialized sub-agents that retrieve and synthesize knowledge across policy proceedings, legal regulations, and license databases.
- The system has been evaluated on a question and answer (Q&A) dataset based on real-world license records and policy proceedings, and has achieved over 80% win rate against strong baselines.
- A new framework for evaluating task-oriented dialogue agents has been proposed, which compiles a workflow specification and per-turn state diff into atomic, schema-grounded criteria and routes each through a cascade of symbolic and encoder/NLI verifiers.
- The framework has been evaluated on four slices spanning MultiWOZ, Schema-Guided Dialogue, and ABCD, and has shown strong performance compared with state-of-the-art baselines.
- Researchers have developed a voice-enabled AI system called Conversation Coach to help practice difficult workplace conversations.
- The system uses a voice-first AI approach to enable managers to rehearse conversations in a realistic spoken format.
- A new method for compressing reasoning chains while preserving answer accuracy and logical coherence has been proposed.
- The Hierarchical Semantic Distillation Network (HSDN) framework combines semantic segmentation, dependency graph construction, dual encoder importance scoring, constrained segment selection, and local boundary rewriting.
- The results show that HSDN achieves 91.0% accuracy with 68.4% compression, outperforming strong compression baselines in overall score and reasoning coherence.
- Researchers have developed a novel Multi-Agent Retrieval-Augmented Generation (RAG) system for spectrum intelligence, enabling autonomous agents to coordinate specialized sub-agents that retrieve and synthesize knowledge across policy proceedings, legal regulations, and license databases.
Sources
- A Stable Aggregation Method for Quantum Federated Learning
- Dr. Claw: An AI Scientist Workspace for Vibe Research
- RestoreBench: Can AI Agents Restore Power Flow Convergence?
- Dependency-Aware Chain-of-Thought Compression for Financial Reasoning
- SpecMind: Enabling Spectrum Intelligence via Multi-Agent Hybrid Retrieval-Augmented Generation
- SAGE: State-Grounded, Abstention-Aware Evaluation of Task-Oriented Dialogue Agents
- Conversation Coach: A Voice-enabled AI System that Helps Practice Difficult Workplace Conversations
- mimeo: Compiling Public Expert Corpora into Agent Skills and Testing What Transfers
- Towards a Belief-Based World Model for LLM Agents
- EGT-KG: Evidence-Grounded Typed KG Retrieval for Practical Scientific QA with Small Language Models
- The Privacy-Hallucination Tradeoff in Differentially Private Language Models
- Validity-Aware Jailbreak Evaluation for Large Language Models
- Wave Function Backpropagation with Explicit Temporal-Interval Dynamics
- CoVer: Conflict-Aware Claim Verification
- When the Algorithm Becomes the Brand Crisis: A Sociotechnical Theory of Distributed Responsibility and Accountable Transparency
- ISO-RAG: Isoperimetric Noise Control for Retrieval-Augmented Generation
- Feedback-Assisted Trust Propagation over Document Relation Graphs for Retrieval-Augmented Generation
- VoiceLongMemEval: Do Assistants Remember How You Sounded?
- Residual Sparsification via Output Importance for Compressing Mixture-of-Experts LLMs
- Consistency Without Alignment: Item-Sensitive Language Models Indistinguishable From Random
- Same Request, Different Boundary: Evaluating Cybersecurity Assistance across Conversational Contexts
- Socrates went Nuclear: Comparing Interaction Strategies for AI systems in a Learning Context using Brain Sensing
- Control-Data Flow Separation: Stable Prompt Optimization in Multi-Agent LLMs
- REVISE: Validity-Guided Recovery for Online Revisions in Agent Workflows
- DramaChain Bench: An End-to-End Benchmark for Short-Drama Generation
- Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search
- SciTrue: Reliable Scientific Claim Validation with Frontier and Open Language Models at the NTCIR SciClaimEval Task
- Drift-Aware LLM Routing with Sparse Contexts and Shared Budgets
- Triple-Bottom-Line Sustainability of Language Models for Edge AI: A Comparison Between SLMs and Quantized LLMs
- Value Over Language Model: Detecting Original Contribution in Writing
- ChatDev 2.0: A No-Code Multi-Agent Platform for Developing Everything
- A Closed-Loop Evaluation of Capability Loss and Recovery in Compressed Driving Policies
- SOVER: Formal Certification of Optimization Reformulations via LLM-Assisted SMT Verification
- Agentic Empirical Asset Pricing: Methodological Foundations
- Escaping Redundant Reasoning: Structure-Aware Search for Inference-Time LLMs
- ContextPipe: Database-Inspired Context Assembly for Long-Horizon Agents
- S^3martCirc: Self-supervised Smart Circuit Discovery
- Automated Tree Knowledge Graph Construction using Ontology Expansion and Retrieval from Vietnamese History Textbooks
- DiagEvo: Diagnosis-Guided Self-Evolution via Hierarchical Error Memory
- When Features Become Instances: Inverted Contrastive Learning for Unsupervised Feature Selection
- StudyBench: Can Self-Evolution Squeeze Textbooks for Olympiad Capability?
- Towards a Reliable and Practical Eval Pipeline
- One Policy, Any Budget: Internalizing Budget-Aware Search via Reinforcement Learning
- AnalysisBank: An Expert Analysis Pattern Library for Financial Report Generation
- Polished but Unresolved: Identifying Late-Stage Pressure States in Long-Horizon Tool-Use Agents
- FLaG: Frequency-Domain Latent-attention Gated Pooling for Token Aggregation
- Towards Generalizable Visually Grounded Exploration of Household Devices
- Verifiable Disaster Storylines and Causal Knowledge Graphs: A Citation-Grounded Pipeline from Heterogeneous Humanitarian Sources
- Reinforcement Learning Enhanced LLM Agents for Complex Vehicle Routing Problems
- Beyond the Clock: Measuring the Value of Adaptive Revision
- FractalNet-Based Heterogeneous Federated Learning for Orbital Edge Intelligence in Satellite Mega-Constellations: A Wildfire Case Study
- Towards reliable multimodal disaster severity assessment through preference optimization and explainable vision-language reasoning
- Denoising Diffusion Generative Models Secretly Calculate Attentions
- CacheBridge: Efficient Cross-Model KV Cache Transfer
- CARE: Contrastive Anchor-based Rubric Evolution for Large Language Model Post-Training
- In-Context Neurofeedback: Can LLMs Control Their Internal Representations through Privileged Access?
- RPCBench: A Benchmark for Proactive Premise Critique in LLM-based Recommendation
- VIBE-Bench: Evaluating Personalized Large Language Models When Profiles Don't Mean Preferences
- Few-Shot Out of Domain Intent Detection with Covariance Corrected Mahalanobis Distance
- CoBRA: Learning Tool-Use Boundaries via Counterfactual Margins
- Figures as Programs: Recursive Generation of Editable Scientific Figures
- Spawn Freely, Act Sparingly: Progressive Risk Vesting for Recursive LLM-Agent Trees
- Data-Driven Persona-Conditioned Agents for A/B Test Simulation
- AgentFactory: Towards Automated Agentic System Design and Optimization
- QILP-0: Constructing Observational Declarative Twins of Quantum Circuits
- WorldBench: Culturally Grounded Benchmark for Multilingual Agents
- User Representation via Cross Multi-source Behavior Pre-training for Mobile Games
- ARISE-RL: Agentic Rubric-Grounded Iterative Self-Evolution with Reinforcement Learning
- Space Generative AI with Solar Energy Harvesting
- Measuring the Behavioral Fidelity of Long-Horizon Human Activity Simulations
- Latent Recurrent Thoughts: Recurrent Refinement of Proposed Latents for Reasoning with Frozen LLMs
- Jailbreaking Text-to-Image Models Through Cracks: Navigating Heterogeneous Safety Filters via Multi-Agent Debate
- FinLifeBench: Exhaustive Life-Event History and Financial-State Reconstruction from Longitudinal Banking Dialogue
- H2Table: Hierarchical Hypergraph-Enhanced Large Language Models for Complex Table Reasoning
- Prompt-Robust Language Models: Which Training Strategies Work?
- Dual Process Motion Planning
- Making Prospective Memory SLM-Shaped: Typed Intention Stores for Small-Model Agents
- Analog-DB: An Agent-First Analog Integrated Circuit Database, From Blocks to Systems
- A Composable Evaluation System for Reproducible Omni-Modal Foundation Model Evaluation
- Automated Event Log Generation from Unstructured Text Using Finetuned LLMs
- LEAP: Likelihood Elicitation and Aggregation for LLM-based Probabilistic Forecasting
- Cheap Verifiers, Large Blind Spots: Measuring the Reliability Cost of Cost-Saving Cascades
- SymFold: Synergizing Evolutionary and Structural Priors for Accurate Protein Inverse Folding
- EDGE: Error Dependency Graph-Guided Multi-Error Attribution in Multi-Agent LLM Systems
- Neuro-Symbolic Geometric Abstraction (NeuSOGA): From Observations to Symbolic Mathematical Representations
- EdiTikZ: Scientific Figure Editing from Revision Trajectories
- Parsing the Stream: A Live Trace Model for Long-Horizon Agents and Their Observers
- Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement
- When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluation
- EvoSCM: Scientific Belief Revision Through Causal Model Evolution and Experimentation
- Can LLMs Discover Scientific Laws in Real and Parallel Worlds?
- Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers
- HyperWorld: Hypergraph-Structured State Serialization Improves Learned Textual World Models
- I-CARE: Analysis of interference-related phenomena in a controllable, diverse and representative unlearning setting for text-to-image models
- Discrete-Time MDP Modeling for Multi-Item Capacitated Lot Sizing with Stochastic Demand Timing
- Incremental Risk Assessment of Progressive Elder Financial Scams via Instruction-Tuned Small Language Models
- Long-Horizon State Tracking in LLMs: Executing MD5 through a Deep Sequence of Dependent Tool Calls
- MiNER: Fine-Tuned Biomedical Natural Language Processing for Malaria Disease Entity Recognition in Clinical Texts
- OpenAgentFlow: Enabling System-Wide Safety Boundaries for Heterogeneous AI Agent Fleets
- SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces
- UI-Venus-2 Technical Report
- EULER: Exploring Underused Links with Evidence-Checked Return for Multi-Agent Mathematical Discovery
- When Prediction Error Is Not Enough: Evaluating Nuisance-Function Prediction for Causal Estimation
- AI Morbidity and Mortality: A Framework for Clinical AI Failure Review
- Different representation learning objectives recover distinct latent structures from the same psychometric data
- Deploying and Evaluating a Smart-Agriculture Agentic Engine for Full-Season Soybean Farm Operations
- Recursive Criticality of AI Self-Improvement
- IMPACT: Attention Is the Interaction Map for Scalable Interaction-Aware World Model Training
- Asymmetries in Spontaneous and Instructed Deception
- LLM-Driven Autonomous Vehicles Inherit Human Driver Biases in Pedestrian Yielding: Results and Implications From A New Benchmark
- ReDeck: Step-Level Render-Grounded Refinement for Document-to-Slide Generation
- AI Should Not Only Be Helpful. It Should Be Contingent. Artificial Intimacy, Sycophancy, and the Future of Social Learning
- ConvDeck: Conversational Paper-to-Slide Generation via Stage-Specific User Feedback
- Learning What to Retain: Gated-Memory Routing for Efficient Collaboration in Multi-Agent LLM Systems
- Invalidation Contracts for Cross-Episode Agent Memory
- Authority Bias in Conversational Search Engines for Academic Paper Recommendation
- Hypotheses-Guided Self Distillation for Continual Personalization
- The Answer Is Not the Argument
- Autoresearch for Marketplace Catalogs: From Legacy Forms to AI-Native Matching
- The Irreversibility Budget: Fleet-Level Risk Accounting and Admission Control for Agent Operating Systems
- The Assistant's Ideal Self
- Human-AI Co-Interpretation for Responsible AI: A Hermeneutic Perspective
- SlideBank: A Persistent Hierarchical Evidence Bank for Consistent Whole-Slide Reasoning
- Vision Is Not Overhead: One-Pass Block Drafting for Lossless Speculative Decoding in Vision-Language Models
Comments
Please log in to post a comment.