Next Reply Prediction X Dataset: Linguistic Discrepancies in Naively Generated Content
Dataset and analysis of linguistic discrepancies when LLMs generate content naively for social science research proxies.
Dataset and analysis of linguistic discrepancies when LLMs generate content naively for social science research proxies.
FUSAR-GPT visual language model for SAR imagery interpretation with spatiotemporal feature embedding.
Unified pushing policy for robotic manipulation using visual prompting with broader applicability across scenarios.
HybridFL framework for federated learning addressing hybrid data distribution in financial crime detection.
Dynamic rollout allocation and advantage modulation for LLM reasoning with verifiable rewards, improving policy optimization.
Independent evaluation of SAP RPT-1 tabular foundation model vs XGBoost/LightGBM on structured enterprise data.
Theoretical analysis of how quantization affects model and data capacities in low-precision high-dimensional linear regression.
LAVIDA uses multimodal LLMs for zero-shot video anomaly detection without labeled anomaly data.
DGPO combines RL fine-tuning with graph diffusion for neural architecture search on directed acyclic graphs.
CORDIC-based vector processing engine for edge AI with runtime-adaptive mixed-precision for IoT applications.
Analysis and solution for preconditioner drift in federated second-order optimizers on non-IID data, improving convergence stability.
Multimodal path planning for multi-agent cooperation using language communication for safe decentralized coordination.
TOPReward uses token probabilities from VLA models as zero-shot reward signals for robot RL training with improved sample efficiency.
Multi-step retrieval method for personalized QA using reasoning to extract relevant personal context beyond direct query-based retrieval.
Taxonomy and empirical analysis of memory systems in LLM agents, evaluating benchmarks, metrics, and performance across model backbones for long-horizon reasoning.
Hierarchical agentic system for automated urban geospatial modification using multimodal reasoning to handle interdependent planning changes.
Policy optimization method bridging GMPO and SAPO using sequence-level importance sampling for improved LLM alignment training.
Soft Adaptive Policy Optimization with smooth gate functions replacing hard clipping for more stable LLM training compared to GRPO.
Framework combining active perception and disentangled representations for continual and few-shot learning without destructive interference.
Theoretical analysis and empirical validation of isotropic Gaussian embeddings for stable deep reinforcement learning under non-stationary environments.
Multi-robot coverage framework integrating Hilbert space-filling curves into DQN and PPO for scalable decentralized exploration and learning.
Empirical study of autonomous coding agents' pull requests on GitHub, analyzing integration outcomes and review collaboration signals using AIDev dataset.
Red-team evaluation of Claude Opus and ChatGPT as security advisors for trusted execution environments, identifying LLM vulnerabilities in security guidance.
Benchmark for time-series reasoning combining general LLM reasoning with domain-specific time-series knowledge through injected knowledge approach.
ContentBench benchmark suite measuring how well low-cost LLMs perform interpretive coding tasks versus humans, tracking cost-accuracy tradeoffs.
Sequential correction algorithm for training physics-informed neural networks more efficiently on partial differential equations.
Evaluates conformal prediction methods for robust uncertainty quantification in EEG clinical diagnosis under distribution shifts.
Interactive browser-based platform for learning federated learning concepts with real-time visualization of heterogeneous data distributions and aggregation algorithms.
Open-source anthropomorphic social robot platform powered by LLMs for accessible AI research in human-robot interaction.
Hierarchical mixture-of-agents architecture for cost-optimized LLM inference routing between small and large models.
Survey on integrating large language models into UAV systems for environmental understanding, swarm coordination, and task reasoning.
Active search algorithm for autonomous agents balancing exploration-exploitation with cost-aware decision making.
Security analysis of LLM-based agentic systems examining runtime attack surfaces, tool invocation risks, and exploitation vectors.
LLM-based text-to-speech system using CTC alignment for low-latency dual-streaming synthesis with improved training sequences.
Comparison of ML approaches (XGBoost, Random Forest, TabNet) for radiation dose estimation in nuclear safety. Physics simulation.
Multimodal sentiment analysis approach using tri-subspace disentanglement for pairwise modality-shared signals.
Heterogeneous Graph Transformer predicting high-potential SMEs from public SBIR funding data. Business intelligence application.
Multimodal learning method addressing semantic misalignment via cross-level collaborative representation fusion.
Machine learning pipeline using satellite imagery to detect looted archaeological sites in Afghanistan. Application of computer vision.
Graph Transformer variant with token-level attention for improved scalability and out-of-distribution generalization on large graphs.
Pedagogically-informed human-AI system for collaborative instructional video generation using cognitive learning theory.
Concept erasure technique for text-to-image diffusion models to prevent harmful content generation via representation misdirection.
Compositional planning approach using pre-trained policy composition for temporal abstraction in agent decision-making.
Particle filtering algorithm for state estimation trained with single-step objectives instead of sequence unrolling. Robotics application.
Study of representational dynamics in minimal continual learning agent with persistent state across executions.
Architecture for embedding carbon-aware governance gates in GenAI-assisted software development workflows.
Open-source MLIR-based compilation stack for deploying Triton kernels and PyTorch models on Qualcomm NPUs.
ML-based detection of malicious pickle-serialized models in repositories like Hugging Face to prevent RCE attacks.
Fault injection framework (MAS-FIRE) for evaluating reliability of LLM-based multi-agent systems under semantic failures.
System-level threat monitoring and reliability challenges for LLM-enabled applications in production environments.