AI
AI News
发布 2026年8月4日
1
将以当前语言复制一段把此列表按国家·公司·技术汇总的提示词。粘贴到 Claude·ChatGPT 等任意 AI 即可。
2026-08-03 ~ 2026-08-04 · AI 主要资讯 80 条。点击标题前往原文。
日期来源关键词标题
- 2026-08-04🇺🇸 Ars Technica AIAIAn AI-supervised remote exam went so badly that 58,000 students must retake it
- 2026-08-03🌐 Import AI (Jack Clark)AIImport AI 467: Self-sustaining AI viruses; pacing AI progress; confusion about AI and creativity
- 2026-08-03🌐 arXiv cs.AIAIOpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems
- 2026-08-03🌐 arXiv cs.AILLMHow Hard Does It Think? Analyzing Step-Aware Reasoning Energy in LLM Chain-of-Thought Trajectories
- 2026-08-03🌐 arXiv cs.AIinferenceIdentifying Informative Environments for Cognition Parameter Inference via Bayesian Experimental Design
- 2026-08-03🌐 arXiv cs.AILLMNeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability
- 2026-08-03🌐 arXiv cs.AILLMMerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations
- 2026-08-03🌐 arXiv cs.AILLMHarnessing the Wisdom of LLM Crowds through Complementarity-Driven Iterative Collaboration
- 2026-08-03🌐 arXiv cs.AIAITool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents
- 2026-08-03🌐 arXiv cs.AILLMModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Models
- 2026-08-03🌐 arXiv cs.AILLMAMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
- 2026-08-03🌐 arXiv cs.AILLMAgentHPOBench: A Benchmark For Evaluating LLM Agents as Sequential Hyperparameter Optimizers
- 2026-08-03🌐 arXiv cs.AILLMThe Formalism Trap: Are LLM-as-a-Judge Evaluators Blinded by Consensus Mimicry under Social Load?
- 2026-08-03🌐 arXiv cs.AItransformerSensitivity Analysis of GRU, LSTM and Transformer Encoder in Classification of Automated Driving Systems
- 2026-08-03🌐 arXiv cs.AILLMGuarantees on Dynamical System Distinguishability for LLM Token Generation
- 2026-08-03🌐 arXiv cs.AILLMMetaphor-Induced Algorithmic Steering: Cross-Domain Procedural Transfer in LLM Code Generation
- 2026-08-03🌐 arXiv cs.AIdeep learningPredicting Steel Fatigue Life from Micrographs Using Physics-Informed Deep Learning
- 2026-08-03🌐 arXiv cs.AIgenerativeDragonCrawl: A Generative, Intent-Based Framework for Scalable Mobile End-to-End Testing
- 2026-08-03🌐 arXiv cs.AILLMBenchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation
- 2026-08-03🌐 arXiv cs.AIdeep learningA Unified Benchmark of Deep Learning Models for Multi-task 3D Brain Tumor Segmentation from Magnetic Resonance Imaging
- 2026-08-03🌐 arXiv cs.AILLMTextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text
- 2026-08-03🌐 arXiv cs.AILLMValidation Evidence in LLM Repair Agents: How Much of What Passes Actually Tests the Bug?
- 2026-08-03🌐 arXiv cs.AILLMTo Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing
- 2026-08-03🌐 arXiv cs.AIAIHuman-LLM Collaborative Inductive Coding for Conceptualizing K-12 Educator AI Use
- 2026-08-03🌐 arXiv cs.AILLMEfficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates
- 2026-08-03🌐 arXiv cs.AILLMA robust association between LLM use and scientific productivity: Assessing stopping-time selection
- 2026-08-03🌐 arXiv cs.AIneuralHERO: History-Enriched Rollout Training for Long-Horizon Autoregressive Neural Operators
- 2026-08-03🌐 arXiv cs.AImachine learningImplicit Machine Learning Force Fields Accelerate Molecular Dynamics Simulations
- 2026-08-03🌐 arXiv cs.AILLMMemory Provenance Laundering in LLM Agents: A Non-Amplification Firewall for Persistent Memory
- 2026-08-03🌐 arXiv cs.AIAISmall Is Enough: Per-User Style Rewriting of AI-Edited Text via LoRA Adapters
- 2026-08-03🌐 arXiv cs.AIAITAVI-TEC: An AI-Based Tool for Procedural Planning of Transcatheter Aortic Valve Implantation
- 2026-08-03🌐 arXiv cs.AILLMCalibratedRubric: Task-Adaptive Rubric Banks for Open-Ended LLM Evaluation
- 2026-08-03🌐 arXiv cs.AItransformerDualDiT: A Conditional Dual-Output Diffusion Transformer for Joint OCT Image and Segmentation Mask Generation
- 2026-08-03🌐 arXiv cs.AIAIFrom Code Review to Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale
- 2026-08-03🌐 arXiv cs.AIAIARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation
- 2026-08-03🌐 arXiv cs.AIinferenceFriendBench: Benchmarking Dyadic Familiarity Inference in Humans and Multimodal Large Language Models
- 2026-08-03🌐 arXiv cs.AILLMWhat Makes a Sale? Simulating End-to-End Seller--Buyer Retail Dynamics with LLM Agents
- 2026-08-03🌐 arXiv cs.AIAISREGym: A Live Benchmark for AI SRE Agents with High-Fidelity Failure Scenarios
- 2026-08-03🌐 arXiv cs.AIinferenceDual-Dimensional Consistency: Balancing Budget and Quality in Adaptive Inference-Time Scaling
- 2026-08-03🌐 arXiv cs.AIAIA Multi-Agent System for Motor Design Optimization via an FEA-AI Hybrid Approach
- 2026-08-03🌐 arXiv cs.AILLMReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning
- 2026-08-03🌐 arXiv cs.AILLMEvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures
- 2026-08-03🌐 arXiv cs.AILLMNeurOWL: An LLM-Based Neural-symbolic Framework for Incomplete OWL Ontology Reasoning
- 2026-08-03🌐 arXiv cs.AIAIDeepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook
- 2026-08-03🌐 arXiv cs.AIneuralMonotone and Separable Set Functions: Characterizations and Neural Models
- 2026-08-03🌐 arXiv cs.AILLMPay for The Second-Best Service: A Game-Theoretic Approach Against Dishonest LLM Providers