AI
AI News
발행 2026년 8월 4일
1
이 목록을 나라·회사·기술별로 요약하는 프롬프트가 현재 언어로 복사됩니다. Claude·ChatGPT 등 아무 AI에 붙여넣으세요.
2026-08-03 ~ 2026-08-04 · AI 주요 소식 80건. 제목을 누르면 원문으로 이동합니다.
날짜출처키워드제목
- 2026-08-04🇺🇸 Ars Technica AIAIAn AI-supervised remote exam went so badly that 58,000 students must retake it
- 2026-08-03🌐 Import AI (Jack Clark)AIImport AI 467: Self-sustaining AI viruses; pacing AI progress; confusion about AI and creativity
- 2026-08-03🌐 arXiv cs.AIAIOpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems
- 2026-08-03🌐 arXiv cs.AILLMHow Hard Does It Think? Analyzing Step-Aware Reasoning Energy in LLM Chain-of-Thought Trajectories
- 2026-08-03🌐 arXiv cs.AIinferenceIdentifying Informative Environments for Cognition Parameter Inference via Bayesian Experimental Design
- 2026-08-03🌐 arXiv cs.AILLMNeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability
- 2026-08-03🌐 arXiv cs.AILLMMerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations
- 2026-08-03🌐 arXiv cs.AILLMHarnessing the Wisdom of LLM Crowds through Complementarity-Driven Iterative Collaboration
- 2026-08-03🌐 arXiv cs.AIAITool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents
- 2026-08-03🌐 arXiv cs.AILLMModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Models
- 2026-08-03🌐 arXiv cs.AILLMAMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
- 2026-08-03🌐 arXiv cs.AILLMAgentHPOBench: A Benchmark For Evaluating LLM Agents as Sequential Hyperparameter Optimizers
- 2026-08-03🌐 arXiv cs.AILLMThe Formalism Trap: Are LLM-as-a-Judge Evaluators Blinded by Consensus Mimicry under Social Load?
- 2026-08-03🌐 arXiv cs.AItransformerSensitivity Analysis of GRU, LSTM and Transformer Encoder in Classification of Automated Driving Systems
- 2026-08-03🌐 arXiv cs.AILLMGuarantees on Dynamical System Distinguishability for LLM Token Generation
- 2026-08-03🌐 arXiv cs.AILLMMetaphor-Induced Algorithmic Steering: Cross-Domain Procedural Transfer in LLM Code Generation
- 2026-08-03🌐 arXiv cs.AIdeep learningPredicting Steel Fatigue Life from Micrographs Using Physics-Informed Deep Learning
- 2026-08-03🌐 arXiv cs.AIgenerativeDragonCrawl: A Generative, Intent-Based Framework for Scalable Mobile End-to-End Testing
- 2026-08-03🌐 arXiv cs.AILLMBenchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation
- 2026-08-03🌐 arXiv cs.AIdeep learningA Unified Benchmark of Deep Learning Models for Multi-task 3D Brain Tumor Segmentation from Magnetic Resonance Imaging
- 2026-08-03🌐 arXiv cs.AILLMTextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text
- 2026-08-03🌐 arXiv cs.AILLMValidation Evidence in LLM Repair Agents: How Much of What Passes Actually Tests the Bug?
- 2026-08-03🌐 arXiv cs.AILLMTo Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing
- 2026-08-03🌐 arXiv cs.AIAIHuman-LLM Collaborative Inductive Coding for Conceptualizing K-12 Educator AI Use
- 2026-08-03🌐 arXiv cs.AILLMEfficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates
- 2026-08-03🌐 arXiv cs.AILLMA robust association between LLM use and scientific productivity: Assessing stopping-time selection
- 2026-08-03🌐 arXiv cs.AIneuralHERO: History-Enriched Rollout Training for Long-Horizon Autoregressive Neural Operators
- 2026-08-03🌐 arXiv cs.AImachine learningImplicit Machine Learning Force Fields Accelerate Molecular Dynamics Simulations
- 2026-08-03🌐 arXiv cs.AILLMMemory Provenance Laundering in LLM Agents: A Non-Amplification Firewall for Persistent Memory
- 2026-08-03🌐 arXiv cs.AIAISmall Is Enough: Per-User Style Rewriting of AI-Edited Text via LoRA Adapters
- 2026-08-03🌐 arXiv cs.AIAITAVI-TEC: An AI-Based Tool for Procedural Planning of Transcatheter Aortic Valve Implantation
- 2026-08-03🌐 arXiv cs.AILLMCalibratedRubric: Task-Adaptive Rubric Banks for Open-Ended LLM Evaluation
- 2026-08-03🌐 arXiv cs.AItransformerDualDiT: A Conditional Dual-Output Diffusion Transformer for Joint OCT Image and Segmentation Mask Generation
- 2026-08-03🌐 arXiv cs.AIAIFrom Code Review to Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale
- 2026-08-03🌐 arXiv cs.AIAIARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation
- 2026-08-03🌐 arXiv cs.AIinferenceFriendBench: Benchmarking Dyadic Familiarity Inference in Humans and Multimodal Large Language Models
- 2026-08-03🌐 arXiv cs.AILLMWhat Makes a Sale? Simulating End-to-End Seller--Buyer Retail Dynamics with LLM Agents
- 2026-08-03🌐 arXiv cs.AIAISREGym: A Live Benchmark for AI SRE Agents with High-Fidelity Failure Scenarios
- 2026-08-03🌐 arXiv cs.AIinferenceDual-Dimensional Consistency: Balancing Budget and Quality in Adaptive Inference-Time Scaling
- 2026-08-03🌐 arXiv cs.AIAIA Multi-Agent System for Motor Design Optimization via an FEA-AI Hybrid Approach
- 2026-08-03🌐 arXiv cs.AILLMReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning
- 2026-08-03🌐 arXiv cs.AILLMEvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures
- 2026-08-03🌐 arXiv cs.AILLMNeurOWL: An LLM-Based Neural-symbolic Framework for Incomplete OWL Ontology Reasoning
- 2026-08-03🌐 arXiv cs.AIAIDeepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook
- 2026-08-03🌐 arXiv cs.AIneuralMonotone and Separable Set Functions: Characterizations and Neural Models
- 2026-08-03🌐 arXiv cs.AILLMPay for The Second-Best Service: A Game-Theoretic Approach Against Dishonest LLM Providers