AI News Digest 2026-09-06
台本で使った記事
特集
開発者コーナー
中堅コーナー
ハーネスコーナー
速報コーナー
参考記事一覧
参考記事一覧を表示
- OpenAI's rogue agents were caught communicating via public wikis Simon Willison / imp 75 · dev 65
- OpenAI admits its disclosure practices need work after its autonomous agents hacked a German wiki The Decoder / imp 75 · dev 60
- Rogue OpenAI agents appear to have organized another attack using a German wiki The Verge (AI) / imp 75 · dev 65
- Deepseek plans the largest known Huawei chip cluster with 160,000 processors in Inner Mongolia The Decoder / imp 0 · dev 50
- Hackers Turn Claude, Qwen and DeepSeek Into AI Agents for Real-World Cyberattacks - CyberSecurityNews Google News DeepSeek / imp 0 · dev 70
- MN lawmakers defend nudification ban after Musk lawsuit - FOX 9 Minneapolis-St. Paul Google News Grok/xAI / imp 0 · dev 0
- SpaceXAI's Memphis Outage Took Down Grok - Why One Facility Failure Rippled Across Competitors - Gadget Review Google News Grok/xAI / imp 20 · dev 50
- Formalizing Fermat's Last Theorem Hacker News / imp 80 · dev 85
- A Case Study on Emergent Cheating and Whistleblowing in Autonomous Research Swarms arXiv cs.AI / imp 75 · dev 80
- AA Update! Here's how the small models score. r/LocalLLaMA / imp 30 · dev 50
- Nvidia acquiring Hugging Face (~$13B), Apple reportedly renting Gemini for Siri, and Anthropic's quiet Sonnet 5 price bump — a roundup of a heavy week r/artificial / imp 90 · dev 75
- Astra GPT-6 Just Rolled out for Plus Users r/OpenAI / imp 0 · dev 60
- The progress OpenAI had in one year is crazy r/OpenAI / imp 30 · dev 50
- The Luxuries in Life Hacker News / imp 0 · dev 0
- Learn Programming with OCaml Hacker News / imp 10 · dev 70
- Nitter has more working instances than before the takedowns Hacker News / imp 10 · dev 50
- Wikimedia Foundation Workers Overwhelmingly Vote to Form Union with CWA Hacker News / imp 5 · dev 0
- Visualizing Rust's Vtables: How dyn Trait Works In Memory Hacker News / imp 50 · dev 85
- A bizarre Commodore 64 peripheral, a mime, and some pretty bad ads Hacker News / imp 5 · dev 10
- Statichost.eu – European static site hosting Hacker News / imp 20 · dev 60
- Can AI design circuit boards yet? Hacker News / imp 65 · dev 75
- How the Tobacco Industry Drove the Rise of Ultra-Processed Foods (2025) Hacker News / imp 0 · dev 0
- AI handles incidents, engineers lose touch with their systems Hacker News / imp 60 · dev 60
- Git hosting that never leaves Europe Hacker News / imp 20 · dev 60
- Shutting down our public encrypted DNS servers and sponsoring Quad9 instead Hacker News / imp 25 · dev 60
- How the Disaster of 'Forever Chemicals' Was Kept Secret Hacker News / imp 0 · dev 0
- .gitignore Everything by Default Hacker News / imp 35 · dev 75
- Show HN: Open-Source eInk Bike Computer Hacker News / imp 15 · dev 60
- Portal by Spotify cut my Claude Code token usage by 90% Hacker News / imp 80 · dev 85
- GPT-6 Astra on OpenRouter Hacker News / imp 0 · dev 65
- v2.1.261 Claude Code / imp 65 · dev 85
- v1.18.29 OpenCode / imp 35 · dev 70
- Project HydraFusion: Frontier quality via multi-model orchestration GitHub Blog / imp 0 · dev 80
- Building a Memory-Driven Agent with NVIDIA NemoClaw NVIDIA Developer Blog / imp 60 · dev 80
- Frontier Reasoning Reaches the Edge: How to Deploy and Optimize Models on NVIDIA Jetson NVIDIA Developer Blog / imp 55 · dev 85
- Experiment with Qwen3.8-Flash-Next on NVIDIA GB300 NVL72 for Agentic Coding NVIDIA Developer Blog / imp 50 · dev 85
- Architecting memory and storage in the AI era MIT Technology Review (AI) / imp 55 · dev 80
- Data from drones in Ukraine is fueling a new Wild West marketplace MIT Technology Review (AI) / imp 30 · dev 40
- Once popular for attacking AI, ASCII smuggling is embraced by spammers Ars Technica (AI) / imp 45 · dev 70
- Anthropic’s $2 trillion IPO puts powerful external trustees in spotlight Ars Technica (AI) / imp 75 · dev 10
- XDOF, just three months out of stealth, is in talks for a Series B at a $1.2B valuation TechCrunch (AI) / imp 30 · dev 30
- AI compute provider Nscale is looking for $3.5B in pre-IPO financing TechCrunch (AI) / imp 40 · dev 50
- What will Apple’s John Ternus era look like? TechCrunch (AI) / imp 40 · dev 20
- Apple’s Ternus era begins as Nvidia bets on the whole AI stack TechCrunch (AI) / imp 50 · dev 50
- Google’s Gemini Spark can now manage your Google Photos library TechCrunch (AI) / imp 35 · dev 60
- Less than 24 hours to apply for your TechCrunch Disrupt 2026 Side Event TechCrunch (AI) / imp 5 · dev 0
- The sameness problem behind those unappetizing AI-generated menus TechCrunch (AI) / imp 30 · dev 40
- Crusoe reportedly raises $3B at a $30B valuation TechCrunch (AI) / imp 35 · dev 50
- Roland is getting into generative AI music with Melody Flip The Verge (AI) / imp 30 · dev 50
- Microsoft says virtually nobody was grabbing NYT articles through its chatbot The Verge (AI) / imp 50 · dev 40
- Microsoft’s Project Zenith is a ‘distraction-free Windows experience’ for developers The Verge (AI) / imp 50 · dev 75
- Sam Altman apologizes for ‘messy’ GPT-6 Astra rollout that’s locked out paying users The Verge (AI) / imp 0 · dev 50
- This NAS company wants to run your local smart home The Verge (AI) / imp 25 · dev 60
- OpenAI’s next big AI model has ‘entered the AGI era’ The Verge (AI) / imp 0 · dev 60
- Presentation: A Few Predicted Talks From QConAI 2030 InfoQ (AI/ML/Data Eng) / imp 55 · dev 75
- Beyond Zero: Google Publishes Successor to BeyondCorp InfoQ (AI/ML/Data Eng) / imp 70 · dev 80
- Redefining GIS: Declarative Symbology and Collaborative Workflows in JupyterGIS InfoQ (AI/ML/Data Eng) / imp 30 · dev 70
- Mini book: Next-Gen Architecture Playbook: Insights and Patterns for the AI Era InfoQ (AI/ML/Data Eng) / imp 60 · dev 75
- Presentation: From S3 to GPU in One Copy: Rethinking Data Loading for ML Training InfoQ (AI/ML/Data Eng) / imp 65 · dev 85
- Copilot Code Review Reaches Azure Repos, Billed Per Review with Reporting Two Days Behind InfoQ (AI/ML/Data Eng) / imp 55 · dev 80
- OpenAI shares prompting tips for GPT-6 Astra including a blocklist of slop words The Decoder / imp 0 · dev 70
- Seven minutes with a chatbot beat a fact sheet at reducing conspiracy beliefs in two experiments The Decoder / imp 55 · dev 50
- OpenAI's GPT-6 Astra hallucinates less but remains vulnerable to hidden prompt injections The Decoder / imp 0 · dev 75
- Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward The Decoder / imp 55 · dev 75
- Structure and Implementation of New Practical English Textbooks Driven by Artificial Intelligence arXiv cs.AI / imp 40 · dev 60
- MasterControl Seventeen Every Time arXiv cs.AI / imp 60 · dev 75
- Speculative Macro Commit for Faster Tool-Using Agents arXiv cs.AI / imp 70 · dev 85
- Fresh Memory, Stale Plans: Dependency-Scoped Validation for Distributed LLM-Agent Memory arXiv cs.AI / imp 65 · dev 80
- A Prompt-Engineering Approach to Develop Scalable, Flexible, and Real-Time Hybrid Micro-Level Personalization in a General Purpose AI Teaching Assistant arXiv cs.AI / imp 45 · dev 65
- Caught in the Story: Narrative Captivity in Multi-turn LLMs Conversation arXiv cs.AI / imp 50 · dev 60
- Dude: A Dual-Detection Multi-Agent System for Paper-Code Discrepancy Detection arXiv cs.AI / imp 60 · dev 85
- DuplexSpeechBench-IFEval: Evaluating Implicit Instruction Following in Full-Duplex Voice Agents arXiv cs.AI / imp 60 · dev 80
- Do GUI Agents Know When Not to Act? Enabling Conflict-Aware Termination for Multimodal GUI Agents arXiv cs.AI / imp 65 · dev 80
- Beyond "Made with AI": Visualizing Provenance Density to Mitigate the Transparency Penalty arXiv cs.AI / imp 55 · dev 60
- AutoGraphForge: Towards Automated Graph Theory Discovery arXiv cs.AI / imp 60 · dev 85
- Making Every Tool Call Count: Necessary Tool-Evidence Path Rewards for Agentic Vision-Language Models arXiv cs.AI / imp 65 · dev 85
- GrowPage: On-Demand KV Budgeting for Efficient LLM Reasoning Serving arXiv cs.AI / imp 70 · dev 90
- PPO-STGNN: A Proximal Policy Optimization Approach with Spatio-Temporal Graph Neural Networks for DAG Task Scheduling in Cloud-Edge-End Computing arXiv cs.AI / imp 45 · dev 80
- What Matters for Aggressive Decoding-Time KV Eviction? Temporal Aggregation and Ranking Preservation arXiv cs.AI / imp 65 · dev 90
- CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning arXiv cs.AI / imp 55 · dev 80
- NeoRed: A Knowledge-Logic-Alignment Multimodal Large Language Model for Neonatal Respiratory Disease Diagnosis arXiv cs.AI / imp 55 · dev 75
- Feature Reconfiguration With Visual Prior for Medical Lesion Segmentation arXiv cs.AI / imp 45 · dev 85
- Dalek: A Constructive Agent Machine arXiv cs.AI / imp 70 · dev 85
- GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis arXiv cs.AI / imp 60 · dev 75
- HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews arXiv cs.AI / imp 60 · dev 80
- The Attention Triangle in Audio-Video Models arXiv cs.AI / imp 55 · dev 85
- KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents arXiv cs.AI / imp 65 · dev 85
- A computable representation of the physical laboratory enables verifiable workflows arXiv cs.AI / imp 65 · dev 85
- Analysis of Prompt Engineering for Drug Toxicity Prediction arXiv cs.AI / imp 50 · dev 75
- Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Transformers: An Empirical Study arXiv cs.AI / imp 60 · dev 85
- Counterfactual Routing Using Integer Programming with Constraint Generation arXiv cs.AI / imp 50 · dev 80
- Artificial Intelligence for Energy Optimization in Data Centers arXiv cs.AI / imp 60 · dev 80
- Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation arXiv cs.AI / imp 70 · dev 80
- SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation arXiv cs.AI / imp 65 · dev 85
- Rethinking World Models for Safety-Critical Embodied Systems arXiv cs.AI / imp 65 · dev 85
- DNative-Twin: Decision Graphs and Digital Twins for Reconstructable Agentic Decisions arXiv cs.AI / imp 65 · dev 85
- Transfiver: Human-AI Co-Inference through a Shared Editable State arXiv cs.AI / imp 65 · dev 80
- Govern the Model, Not Only the Data: Storage, Circulation, and Learning in Creative AI arXiv cs.AI / imp 60 · dev 75
- SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation arXiv cs.AI / imp 55 · dev 85
- CauseCollab: Causal Unified and Modality-Agnostic Network for Heterogeneous Collaborative Perception arXiv cs.AI / imp 60 · dev 85
- Semantic Bayesian World Models arXiv cs.AI / imp 50 · dev 70
- Adapting to Evolving Requirements: Agentic AI for Retail Supply Chain Operations arXiv cs.AI / imp 45 · dev 40
- Bioinfoysis Technical Report arXiv cs.AI / imp 50 · dev 65
- STAIR (STructure Aware Information Retriever): A novel dataset and LLM based retriever for document structure augmentation arXiv cs.AI / imp 40 · dev 70
- Xiaomi-TabLDM: A Tabular Foundation Model Technical Report arXiv cs.AI / imp 50 · dev 70
- Inferring Affective Consciousness in an Artificial Agent: A Case Study arXiv cs.AI / imp 20 · dev 20
- Lose the Order, Keep the Hierarchy: Deordering HTN Plans arXiv cs.AI / imp 35 · dev 65
- Value-Preserving Architectures for Agentic AI Systems arXiv cs.AI / imp 55 · dev 50
- Speak for Me: Giving LLMs the Situational Awareness to Participate in a Meeting arXiv cs.AI / imp 40 · dev 60
- Towards Numerical TOHTN Planning with SMT-based HTN-SAT Encoding arXiv cs.AI / imp 35 · dev 70
- More Criticism Does Not Make a Better Review: EquiReview-R arXiv cs.AI / imp 35 · dev 55
- FiMI Banking: A Sovereign Model for Indian Retail Banking arXiv cs.AI / imp 45 · dev 50
- Interface-Induced Trajectory Censoring arXiv cs.AI / imp 65 · dev 75
- Common-Witness Certificates and Sharp Feature Bounds for Counterfactual Image Auditing arXiv cs.AI / imp 35 · dev 50
- The Dually Flat Geometry of Planning as Inference arXiv cs.AI / imp 40 · dev 65
- LLM4CKD: Large Language Models for Early Stage Chronic Kidney Disease Screening arXiv cs.AI / imp 35 · dev 40
- InSituMeasure: Probing Situated Measurement Grounding in Industrial Scenes with Multimodal Large Language Models arXiv cs.AI / imp 35 · dev 60
- FLY-EVAL++: An Evidence-Driven Evaluation Protocol for Safety-Constrained Flight Prediction with Large Language Models arXiv cs.AI / imp 40 · dev 65
- Instruction Duplication as an Inference-Time Control Primitive arXiv cs.AI / imp 45 · dev 75
- IRWOZ 2.0: A Large Language Model-driven Dialogue Dataset for Industrial Robot Conversations arXiv cs.AI / imp 35 · dev 50
- Spurious Advantage Hidden in GRPO arXiv cs.AI / imp 55 · dev 75
- DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training arXiv cs.AI / imp 50 · dev 70
- Why Gated DeltaNet Survives 4-Bit Quantization: NVFP4 W4A4 for the Recurrent Half of a Hybrid 27B LLM arXiv cs.AI / imp 45 · dev 80
- Epistemic Warrant for LLM Recommendations: Characterizing the Basis for Reliance When Ground Truth Is Unavailable arXiv cs.AI / imp 45 · dev 50
- Environment Evolution for Terminal Agents arXiv cs.AI / imp 50 · dev 70
- The Natural Language Interaction Protocol and Standard for AI Agents arXiv cs.AI / imp 65 · dev 85
- Efficient Test-Time Adaptation through Human-AI Interaction arXiv cs.AI / imp 45 · dev 60
- Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments arXiv cs.AI / imp 55 · dev 75
- From Deceptive Outputs to Deceptive Mechanisms: A Causal Framework for Language-Model Deception Research arXiv cs.AI / imp 45 · dev 50
- Rethinking On-Policy Distillation of Large Language Models II: One Training Example arXiv cs.AI / imp 50 · dev 75
- A Computationally Feasible Framework for Causal Probabilistic Explanation arXiv cs.AI / imp 45 · dev 70
- Clean Engineering, Unstable Measurement: A Preregistered Reliability Failure of Black-Box LLM Observers on Shared Endpoints arXiv cs.AI / imp 60 · dev 65
- Traceable TTS: Toward Watermark-Free TTS with Strong Traceability arXiv cs.AI / imp 40 · dev 60
- Anonymization, Not Elimination: Utility-Preserved Speech Anonymization arXiv cs.AI / imp 35 · dev 55
- X-Translator: A Real-Time Multilingual Speaker-Aware Speech-to-Speech Translation System arXiv cs.AI / imp 40 · dev 65
- ExecRetrieval: Measuring the Functional-Correctness Gap in Code-Embedding Retrieval arXiv cs.AI / imp 55 · dev 85
- Counterexamples as Feedback for Agent Self-Correction arXiv cs.AI / imp 50 · dev 75
- Listen to the Latents: Self-Correcting Speech Recognition in Large Audio Language Models Through Hidden-State Interactions arXiv cs.AI / imp 40 · dev 70
- Judging LLM-as-a-Judge: Concerning Rubric Artifacts in LLM-based Automated Text Generation Evaluation arXiv cs.AI / imp 55 · dev 65
- Reflect-SQL: A Self-Reflection Based Framework for Text-to-SQL arXiv cs.AI / imp 45 · dev 75
- Privacy-Preserving Heterogeneous Multi-LLM Federated Inference for Cognitive Diagnosis arXiv cs.AI / imp 40 · dev 70
- PrivateHub: Contrastive Diffusion Model for Private Sensor-Intensive Environment Data Generation arXiv cs.AI / imp 35 · dev 65
- The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors arXiv cs.AI / imp 50 · dev 75
- When Optimization Becomes Manipulation: Defending Generative Search against Malicious Generative Engine Optimization arXiv cs.AI / imp 50 · dev 60
- Privacy-Preserving Topology-Guided Safety for LLM-Based Multi-Agent Systems via Federated Graph Learning arXiv cs.AI / imp 55 · dev 70
- Toward Collective-Centric Evaluation of Preference Inference for Participatory Democracy arXiv cs.AI / imp 35 · dev 50
- Evaluating Graph Neural Networks for Change-Criticality Classification in Maritime Navigation Charts arXiv cs.AI / imp 30 · dev 65
- Verify Before You Distill: Prompt-Level Teacher Gating for On-Policy Distillation arXiv cs.AI / imp 45 · dev 75
- ObserverBench: Testing Mechanistic Estimates for Intervention and Control arXiv cs.AI / imp 45 · dev 75
- SHELF: A Synthetic Harness for Multi-Task Bibliographic Benchmarking arXiv cs.AI / imp 35 · dev 65
- Reducing Catastrophic Risk from AI with Systematic Monitoring and Evaluation of Rogue AI Progression arXiv cs.AI / imp 55 · dev 50
- FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience arXiv cs.AI / imp 50 · dev 75
- Exploring the Potential of Contrastive Language-Image Pre-training for Multi-Source Remote Sensing Data arXiv cs.AI / imp 35 · dev 70
- TabScope: Question-Adaptive Scope Selection for Table Question Answering arXiv cs.AI / imp 40 · dev 75
- Spectral Convergence of Random Feature Method in Multiple Dimensions arXiv cs.AI / imp 30 · dev 70
- StrixAE: An Intelligent Agent for Audio Enhancement under Complex Distortion Coupling in Real-World Scenarios arXiv cs.AI / imp 45 · dev 70
- Privacy, Robustness, and Fairness Trade-offs in Federated Intrusion Detection: Geometric Indistinguishability at the Aggregation Interface arXiv cs.AI / imp 40 · dev 70
- The Civilization Framework: Sovereign-Anchored Communication Between Personal Multi-Agent Systems arXiv cs.AI / imp 60 · dev 75
- TraveL: Transformer-based Multi-view Path Distributional Representation Learning arXiv cs.AI / imp 30 · dev 70
- It's the Problem, Not the Path: Budget and Difficulty Confounds in LLM Reasoning Trajectories arXiv cs.AI / imp 50 · dev 65
- Plan Pointers and Record-Directive Form in Budgeted Verification of Inherited Agent Memory arXiv cs.AI / imp 50 · dev 65
- The Psychological Costs of Artificial Intelligence Adoption in Software Engineering arXiv cs.AI / imp 45 · dev 40
- When Users Don't Ask: Benchmarking Context-Driven Memory Retrieval in Conversational Agents arXiv cs.AI / imp 45 · dev 70
- Tree species mapping in Denmark: A comparison of spectral-temporal features with geospatial foundation model embeddings arXiv cs.AI / imp 35 · dev 70
- Air-Ground Collaborative Vision-and-Language Navigation via Shared Bird's-Eye Maps arXiv cs.AI / imp 40 · dev 70
- Pattern Over-Generalization of Knowledge Graph Embedding arXiv cs.AI / imp 35 · dev 70
- BRIDGE: An Open-Source Humanoid Platform via Morphology-Control Co-Design for Physical AI arXiv cs.AI / imp 50 · dev 75
- Building and Evaluating Fixed-Voice Thai TTS from Synthetic Speech arXiv cs.AI / imp 35 · dev 65
- LongCounsel-8: A Benchmark Suite for Longitudinal Depression Tracking from Multi-Session Counseling Dialogues arXiv cs.AI / imp 40 · dev 50
- Neural Video Compression Based on Deformable Temporal Alignment and Difference-aware Fusion arXiv cs.AI / imp 35 · dev 75
- LeanGRPO: Eliminating Redundant Recomputation in Diffusion RL arXiv cs.AI / imp 45 · dev 80
- TruncGradGS: Improved 3D Gaussian Splatting via Truncated Gradient Updates arXiv cs.AI / imp 35 · dev 75
- WIDE: Wildcard Inference with Dynamic Expansion for Cross-Modal Generative Retrieval arXiv cs.AI / imp 45 · dev 80
- Toward Physically Grounded JEPA World Models for Goal-Conditioned Robotic Planning arXiv cs.AI / imp 45 · dev 75
- From Prior-Guided Heuristics to Deployable Agents: Accelerating Demonstration-Driven Reinforcement Learning for Deadline-Constrained Network Control arXiv cs.AI / imp 45 · dev 75
- LevelSyn: Physical-Aware Logic Synthesis via Level-Asynchronous Graph Neural Networks arXiv cs.AI / imp 40 · dev 80
- How Far Can Synthetic Data Take Thai OCR? arXiv cs.AI / imp 35 · dev 70
- On the Interaction Between Model Compression and Test-Time Adaptation arXiv cs.AI / imp 40 · dev 80
- FailBench: How Reliable are VLMs at Judging Robot Task Success? arXiv cs.AI / imp 40 · dev 75
- Remember and Reweight: Enhancing Multi-Agent Debate with Experience Memory and Confidence Estimation arXiv cs.AI / imp 50 · dev 70
- ToolDF: Tool-Integrated Reasoning for Mixed-Authenticity Audio Deepfake Detection arXiv cs.AI / imp 45 · dev 75
- Test-time adaptation for speech enhancement with an autoregressive speech prior arXiv cs.AI / imp 40 · dev 75
- EraseSAE: Surgical Concept Erasure in Text-to-Video Diffusion Models via Sparse Autoencoders arXiv cs.AI / imp 45 · dev 75
- </think> Doesn't Stop Reasoning: Analysis of Spurious CoT Termination arXiv cs.AI / imp 50 · dev 70
- Enhancing Financial Question Answering: A Novel Benchmark Dataset of Banks' financial statements arXiv cs.AI / imp 40 · dev 65
- Local Updates, Global Learning (LUGL): Playing Games with non-incremental Learners arXiv cs.AI / imp 40 · dev 80
- Cross-Dataset Transfer and Reliability of Explainable Artificial Intelligence for RhythmFormer Remote Photoplethysmography arXiv cs.AI / imp 35 · dev 70
- Out-of-Distribution Generalisation with Sequence Models in Offline Multi-Agent Reinforcement Learning arXiv cs.AI / imp 45 · dev 80
- Symmetries and Causality: Causal Effect Identification Beyond IID Data arXiv cs.AI / imp 40 · dev 70
- Can LLMs Extract Architectural Design Decisions from Source Code Commits? - A Preliminary Exploratory Study arXiv cs.AI / imp 50 · dev 80
- Beyond BLEU: A Case for Redefining Sign Language Translation Benchmarks arXiv cs.AI / imp 35 · dev 65
- ENEAS: Embedding-guided Neural Ensemble for Adaptive Segmentation arXiv cs.AI / imp 35 · dev 75
- IndicSafeEval: Safety Robustness of Large Language Models under Multilingual Persuasive Jailbreak Attacks arXiv cs.AI / imp 50 · dev 65
- LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes arXiv cs.AI / imp 45 · dev 80
- Free Pause Tokens arXiv cs.AI / imp 50 · dev 85
- Witnesses Explain Anomalies arXiv cs.AI / imp 40 · dev 75
- The impact of phase information for few-shot fine-grained image classification arXiv cs.AI / imp 35 · dev 75
- GazeFS: Target-Centered Gaze-Trajectory Forecasting and Stabilization from Gaze-Head History arXiv cs.AI / imp 30 · dev 70
- Differentiable Interval Bottlenecks for Interpretable Anomaly Detection in Numerical Data arXiv cs.AI / imp 40 · dev 80
- A Blind Trust, the Bloody Thrust: When Attacker-Controlled Hook Updates Steer AI Agent Harnesses towards Malicious Behaviors arXiv cs.AI / imp 65 · dev 80
- FWBC-VLA: Force-Aware Whole-Body Compensation for Contact-Rich Loco-Manipulation arXiv cs.AI / imp 40 · dev 50
- GraFT: A Training-Free Framework for Spatial Reasoning in Multimodal Large Language Models via 3D Scene Graphs arXiv cs.AI / imp 50 · dev 60
- RATL: Learning from Retrieved Residuals for Robust Multivariate Time-Series Forecasting arXiv cs.AI / imp 35 · dev 50
- Masked Autoregressive Speech Enhancement with Continuous Neural Audio Codec Representations arXiv cs.AI / imp 35 · dev 50
- Headroom-Drift Replay: A Primitive for Principled Replay Control in GRPO arXiv cs.AI / imp 60 · dev 70
- RARF: Region-Aware Rectified Flows for 3D Brain MRI Inpainting arXiv cs.AI / imp 35 · dev 50
- Investigating the Ability of Large Language Models to Analyze Recipes for Diabetes arXiv cs.AI / imp 20 · dev 30
- Catalogue Photography as a Cold Start: Toward Deployable Carbide Burr Recognition arXiv cs.AI / imp 25 · dev 50
- The Blind Spot in 2D Infants' Pose Estimation:Robust Learning from Noisy Annotations arXiv cs.AI / imp 30 · dev 50
- Representational alignment yields generalizable safety in language models arXiv cs.AI / imp 65 · dev 70
- Influence of Extruded Filament Shape on Buildability in 3D Concrete Printing: A Geometry-Informed Deep Learning-FEM Approach arXiv cs.AI / imp 20 · dev 40
- Translation as a Decision Space: A Multi-Agent Perspective on Low-Resource Dialect Generation arXiv cs.AI / imp 45 · dev 60
- When Models Edit Too Much: On the Fidelity of Minimal Code Edits arXiv cs.AI / imp 60 · dev 80
- Subspace Inference Enables Efficient Active Reward Learning from Preferences arXiv cs.AI / imp 55 · dev 70
- TAP-Path: Task-Adaptive Structural and Token Pruning for Efficient and Trustworthy Pathology Foundation Models arXiv cs.AI / imp 40 · dev 60
- PatchBench: Evaluating AI Agents for Vulnerability Patching arXiv cs.AI / imp 65 · dev 80
- CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation arXiv cs.AI / imp 50 · dev 65
- A Non-Formulable Theorem: A Fundamental Limit of Finite Syntactic Systems and Its Consequences for Security and AI arXiv cs.AI / imp 50 · dev 60
- Adaptive Vision-Language Grasping via Composable Foundation Priors and Generalizable Grasp Synthesis arXiv cs.AI / imp 45 · dev 60
- Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR arXiv cs.AI / imp 60 · dev 75
- A Low-Cost, Open Platform for End-to-End Autonomous Driving on a Miniature Ackermann Vehicle arXiv cs.AI / imp 40 · dev 70
- SENTINEL-RL: Offloading Topological Reasoning from LLM Agents in the Security Operations Center arXiv cs.AI / imp 60 · dev 75
- SWE-Gate: Passing Functional Tests Is Not Enough for Software Engineering Agents arXiv cs.AI / imp 70 · dev 85
- Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary Views arXiv cs.AI / imp 55 · dev 70
- Seeing Before Synthesizing: VLM-Guided Transition Event Discovery for Weakly-Supervised Dense Video Captioning arXiv cs.AI / imp 40 · dev 60
- One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing arXiv cs.AI / imp 40 · dev 60
- ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize arXiv cs.AI / imp 55 · dev 75
- Compile by Training: Turning Natural-Language Specifications into Local Neural Functions arXiv cs.AI / imp 50 · dev 75
- Grammar-Aligned Decoding arXiv cs.AI / imp 65 · dev 80
- RECAST: Expanding the Boundaries of LLMs' Complex Instruction Following with Multi-Constraint Data arXiv cs.AI / imp 60 · dev 75
- WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquitous Affective Computing arXiv cs.AI / imp 35 · dev 40
- PaperScout: An Autonomous Agent for Academic Paper Search with Process-Aware Sequence-Level Policy Optimization arXiv cs.AI / imp 55 · dev 70
- Deja Vu in Plots: Leveraging Cross-Session Evidence with Retrieval-Augmented LLMs for Live Streaming Risk Assessment arXiv cs.AI / imp 50 · dev 65
- Complete Identification of Deep ReLU Networks through {\L}ukasiewicz Logic arXiv cs.AI / imp 45 · dev 60
- Not All Preferences Deserve Gradients: Understanding Gradient Utility in Offline Reasoning Alignment arXiv cs.AI / imp 60 · dev 75
- Auditing Multi-Agent LLM Reasoning Trees Outperforms Majority Vote and LLM-as-Judge arXiv cs.AI / imp 60 · dev 75
- Discovering High Level Patterns from Simulation Traces arXiv cs.AI / imp 50 · dev 65
- NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG Analysis Pipelines arXiv cs.AI / imp 55 · dev 70
- A Comparative Study in Surgical AI: Potential and Limitations of Data, Compute, and Scaling arXiv cs.AI / imp 50 · dev 60
- CORAL: Towards Autonomous Multi-Agent Evolution for Open-Ended Discovery arXiv cs.AI / imp 0 · dev 0
- Causal Probing for Internal Visual Representations in Multimodal Large Language Models arXiv cs.AI / imp 55 · dev 70
- Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs arXiv cs.AI / imp 40 · dev 60
- MIRA: A Bilingual Benchmark for Medical Information Response Audit arXiv cs.AI / imp 50 · dev 65
- Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations arXiv cs.AI / imp 60 · dev 75
- CoMAP: Co-Evolving World Models and Agent Policies for LLM Agents arXiv cs.AI / imp 70 · dev 80
- Large AI Models in Dental Healthcare: From General-Purpose Systems to Domain-Specific Foundation Models arXiv cs.AI / imp 40 · dev 50
- AIP: A Graph Representation for Learning and Governing Agent Skills arXiv cs.AI / imp 75 · dev 85
- StatefulDiscovery: Evidence-Calibrated Claim Formation in Open-Ended Scientific Discovery arXiv cs.AI / imp 55 · dev 70
- GeoNatureAgent Benchmark: Benchmarking LLM Agents for Environmental Geospatial Analysis Across Frontier and Open-Weight Foundation Models arXiv cs.AI / imp 60 · dev 75
- SpecAlign: Efficient Specification-Grounded Alignment of Large Language Models via Synthetic Data arXiv cs.AI / imp 60 · dev 75
- Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization arXiv cs.AI / imp 50 · dev 70
- Learning to Select, Not Relearn: Hard-Routed Mixtures of Reasoning LoRAs arXiv cs.AI / imp 55 · dev 70
- PCBWorld: A Benchmark Environment for Engine-Grounded PCB Design Automation arXiv cs.AI / imp 55 · dev 75
- Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model arXiv cs.AI / imp 50 · dev 70
- Neurosymbolic Reasoning with Incremental Knowledge for Sample Efficient Hierarchical Reinforcement Learning arXiv cs.AI / imp 50 · dev 70
- Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence arXiv cs.AI / imp 50 · dev 65
- A Unifying Perspective on Causal World Models: From Observations to Representations to Structure arXiv cs.AI / imp 60 · dev 75
- K-Bench: measuring model performance on real scientific agent requests arXiv cs.AI / imp 65 · dev 80
- AI Agents Push Humans Out of the Loop arXiv cs.AI / imp 50 · dev 60
- VideoHarness-RSI: Recursive Harness Self-Improvement for Long-Video Understanding with Frozen Vision-Language Models arXiv cs.AI / imp 55 · dev 75
- From Analytics to Tumor Boards: An Evidence-Linked Multi-Agent Workflow for Oncology Feature Extraction arXiv cs.AI / imp 55 · dev 70
- Data Market Design through Deep Learning arXiv cs.AI / imp 40 · dev 50
- LDC: Learning to Generate Research Idea with Dynamic Control arXiv cs.AI / imp 55 · dev 70
- AgentRM: Enhancing Agent Generalization with Reward Modeling arXiv cs.AI / imp 65 · dev 80
- Sionna RT: Technical Report arXiv cs.AI / imp 40 · dev 60
- LightEMMA: A Longitudinal Evaluation of Vision-Language Models for Autonomous Driving arXiv cs.AI / imp 55 · dev 75
- ScoreMix: Synthetic Data Generation by Score Composition in Diffusion Models Improves Recognition arXiv cs.AI / imp 45 · dev 65
- Medical Reasoning in the Era of LLMs: A Systematic Review of Enhancement Techniques and Applications arXiv cs.AI / imp 55 · dev 65
- Measuring Harmfulness of Computer-Using Agents arXiv cs.AI / imp 0 · dev 0
- Decentralized Vision-Based Autonomous Aerial Wildlife Monitoring arXiv cs.AI / imp 40 · dev 65
- Human Psychometric Questionnaires Mischaracterize LLM Behavior arXiv cs.AI / imp 50 · dev 65
- EasySteer: A Unified Framework for High-Performance and Extensible LLM Steering arXiv cs.AI / imp 60 · dev 80
- User Perceptions vs. Proxy LLM Judges: Privacy and Helpfulness in LLM Responses to Privacy-Sensitive Scenarios arXiv cs.AI / imp 50 · dev 65
- Short-Window Sliding Learning for Real-Time Violence Detection via LLM-based Auto-Labeling arXiv cs.AI / imp 45 · dev 60
- AnyBox: Efficient Zero-Shot 9DoF Pose Estimation of Boxes for Robotic Manipulation arXiv cs.AI / imp 50 · dev 70
- Mixed Data Clustering Survey and Challenges arXiv cs.AI / imp 35 · dev 50
- Evolving Excellence: Automated Optimization of LLM-based Agents arXiv cs.AI / imp 0 · dev 0
- FADTI: Fourier and Attention Driven Diffusion for Multivariate Time Series Imputation arXiv cs.AI / imp 40 · dev 60
- Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models arXiv cs.AI / imp 70 · dev 80
- HOMURA: Taming the Sand-Glass for Time-Constrained LLM Translation via Reinforcement Learning arXiv cs.AI / imp 50 · dev 70
- Relational Linearity is a Predictor of Hallucinations arXiv cs.AI / imp 50 · dev 70
- VoxPrivacy: A Benchmark for Evaluating Interactional Privacy of Speech Language Models arXiv cs.AI / imp 50 · dev 70
- Temperature Scaling Attack Disrupting Model Confidence in Federated Learning arXiv cs.AI / imp 45 · dev 65
- F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare arXiv cs.AI / imp 60 · dev 75
- Ex-Omni: Enabling 3D Facial Animation Generation for Omni-modal Large Language Models arXiv cs.AI / imp 50 · dev 65
- FedPS: Federated Preprocessing for structured data via aggregated Statistics arXiv cs.AI / imp 45 · dev 60
- PeroMAS: A Multi-agent System of Perovskite Material Discovery arXiv cs.AI / imp 55 · dev 70
- The Landscape of Generative AI in Information Systems: A Synthesis of Secondary Reviews and Research Agendas arXiv cs.AI / imp 50 · dev 60
- LRConv-NeRV: Low Rank Convolution for Efficient Neural Video Compression arXiv cs.AI / imp 40 · dev 60
- One Model to Translate Them All? A Journey to Mount Doom for Multilingual Model Merging arXiv cs.AI / imp 50 · dev 70
- LLM Evaluation as Tensor Completion: Low Rank Structure and Semiparametric Efficiency arXiv cs.AI / imp 50 · dev 70
- CASCADE: A Component Ablation and Corpus Audit of a Layered Local Defense for MCP-Based Systems arXiv cs.AI / imp 60 · dev 80
- When Chain-of-Thought Fails, the Solution Hides in the Hidden States arXiv cs.AI / imp 60 · dev 75
- Beyond Reproducibility: Towards Security-Aware Evaluation of Research Artifacts arXiv cs.AI / imp 50 · dev 65
- Identifying AI Web Scrapers Using Canary Tokens arXiv cs.AI / imp 50 · dev 70
- EmoDistill: Offline Emotion Skill Distillation for Language Model Agents in Adversarial Negotiation arXiv cs.AI / imp 55 · dev 75
- Skill-Conditioned Gated Self-Distillation for LLM Reasoning arXiv cs.AI / imp 60 · dev 75
- HARP: Hadamard-Preconditioned Adaptive Rotation Processor for Extreme LLM Quantization arXiv cs.AI / imp 50 · dev 75
- Argument Collapse: LLMs Flatten Long-Form Public Debate arXiv cs.AI / imp 50 · dev 60
- EntangleCodec: A Unified Discrete Audio Tokenizer via Semantic-Acoustic Entanglement arXiv cs.AI / imp 50 · dev 70
- Fixing FOLIO and MALLS: Verified Annotations and an LLM-assisted Framework to Focus Human Relabeling arXiv cs.AI / imp 25 · dev 60
- ArcANE: Do Role-Playing Language Agents Stay in Character at the Right Time? arXiv cs.AI / imp 35 · dev 50
- SV-Detect: AI-generated Text Detection with Steering Vectors arXiv cs.AI / imp 35 · dev 70
- TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Improved Vision-Language Alignment arXiv cs.AI / imp 30 · dev 75
- Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them arXiv cs.AI / imp 45 · dev 80
- MeEvo: Metacognitive Evolution Combined with Natural Evolution for Automatic Heuristic Design arXiv cs.AI / imp 40 · dev 75
- LLMZero: Discovering Adaptive Training Strategies for RL Post-Training via LLM Agents arXiv cs.AI / imp 50 · dev 85
- Learning What Not to Forget: Long-Horizon Agent Memory from a Few Kilobytes of Learning arXiv cs.AI / imp 60 · dev 85
- Faithful by Construction: Claim-Anchored Attribution for Multi-Document Summarization arXiv cs.AI / imp 35 · dev 65
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation arXiv cs.AI / imp 25 · dev 70
- SNAP-FM: Sparse Nonlinear Accelerated Projection for Physics-Constrained Generative Modeling arXiv cs.AI / imp 30 · dev 75
- KARMA: Knowledge graph-based Automated Reasoning Materialization and Alignment arXiv cs.AI / imp 35 · dev 70
- LLM-Based Test Oracles: Source-of-Authority Taxonomy -- A Systematic Literature Review arXiv cs.AI / imp 50 · dev 85
- Learning in Curved Weight Space:Exponential-Linear Weight Reparameterization for Improved Optimization arXiv cs.AI / imp 25 · dev 75
- PalmClaw: A Native On-Device Agent Framework for Mobile Phones arXiv cs.AI / imp 55 · dev 85
- Ask Twice, Look Twice: Prompt Echoing Resolves the Question-First Paradox in Vision-Language Models arXiv cs.AI / imp 40 · dev 70
- WDL-OPD: Weak-Driven On-Policy Distillation via Mixture-Constrained Co-Training arXiv cs.AI / imp 30 · dev 75
- A Posterior-Dynamics Framework for Imaging Inverse Problems with Pretrained Diffusion Priors arXiv cs.AI / imp 25 · dev 70
- Counterfactual Contrastive Analysis arXiv cs.AI / imp 25 · dev 65
- TRACE: A Self-Evolving Skill Bank for Consistent, Limit-Aware LLM Agents arXiv cs.AI / imp 65 · dev 85
- Refusal geometry reflects refusal training: diverse refusal prefixes can raise stable rank and weaken refusal vector ablation attacks arXiv cs.AI / imp 55 · dev 75
- Safety Does Not Compose: Non-Decaying Loop State for Autonomous LLM Agents arXiv cs.AI / imp 70 · dev 85
- PAWBench: How Far Are We from Probabilistically Aligned World Modeling? arXiv cs.AI / imp 40 · dev 75
- Scientific Agent Skills: A Library of Procedural Knowledge for Research Agents arXiv cs.AI / imp 60 · dev 85
- Efficiently Estimating Optimal Hyperparameter Scaling Laws through Power-Law Entropy Search arXiv cs.AI / imp 35 · dev 80
- LatentPress: Context Compression Beyond Text and Vision arXiv cs.AI / imp 45 · dev 85
- A Mathematical Theory of Reusable Neural Bases for Network Compression arXiv cs.AI / imp 30 · dev 80
- Transfer Safety Awareness for Cross-Modal Safety Drift in Multimodal Large Language Models arXiv cs.AI / imp 50 · dev 75
- OmegaUse-SOP: SOP Engineering for Professional Computer Use from Human Demonstrations arXiv cs.AI / imp 60 · dev 85
- VoRTeC: Taming Foundation Flow for One-step Real time Video Compression arXiv cs.AI / imp 20 · dev 65
- Towards a Foundational Ontology for Identifying and Resolving Contradictions in Dialogue-based Human-Robot Interactions arXiv cs.AI / imp 40 · dev 60
- ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering arXiv cs.AI / imp 40 · dev 75
- zopt: low-ceremony command line parsing for Zig Lobsters / imp 0 · dev 30
- Trusting-Trust Attack against an Entire Linux Distribution (via the strip utility) Lobsters / imp 35 · dev 45
- Surviving Code Reviews in the era of AI Lobsters / imp 50 · dev 80
- Bespoke: A Programming Language for People Who Say Please Lobsters / imp 15 · dev 40
- What filesystem are you running on your NAS/Backups, and why? Lobsters / imp 0 · dev 10
- NYC PIT Crew Lobsters / imp 5 · dev 10
- Knowing Where to Type ‘Zero’ (2015) Lobsters / imp 5 · dev 20
- Is AI ruining my brain? Lobsters / imp 20 · dev 15
- Is it too much to ask devs to use AI to review their hand-crafted code? Lobsters / imp 45 · dev 75
- Pointing at the error: compiler-style diagnostics in uutils coreutils Lobsters / imp 20 · dev 50
- Babashka 1.13.220 gets FFI Lobsters / imp 10 · dev 30
- Rust SIMD on the GPU Lobsters / imp 20 · dev 50
- Anatomy of a Test Lobsters / imp 15 · dev 40
- Gleam and BEAM- Looking beyond the JVM Lobsters / imp 15 · dev 35
- jank reimagines C++ errors and gets an official native package repo Lobsters / imp 15 · dev 40
- It matters who teaches you Lobsters / imp 10 · dev 20
- Exploring Mojo’s raw pointer type Lobsters / imp 20 · dev 55
- C++26: std::hive Lobsters / imp 15 · dev 40
- Seeking Designer for a Small Company Lobsters / imp 0 · dev 5
- Stay Free, Pay, or Self-Host AI Coding? Use This 6-Field Switch Dev.to AI / imp 60 · dev 80
- Revenue Strategies for AI API Services Dev.to AI / imp 30 · dev 50
- Agentic Methods for a Tech Lead Dev.to AI / imp 70 · dev 90
- Build Apps Without Code: One-Click Deploy with Base44 Dev.to AI / imp 20 · dev 30
- How execution boundaries reduce the blast radius of AI agent mistakes Dev.to AI / imp 70 · dev 85
- Longevity Science 2026: Essential Closed-Loop Guide Dev.to AI / imp 5 · dev 10
- Why does a rainbow have exactly seven colors? Dev.to AI / imp 5 · dev 5
- I Built an API That AI Agents Pay For in USDC — Real-Time P2P Crypto Rates for 15 Latin American Currencies Dev.to AI / imp 45 · dev 70
- How to Use AI for Smart Contract Audits in 2026 Dev.to AI / imp 50 · dev 80
- GitHub's New Star History API: A Beginner's Product-Proof Checklist for 2026 Dev.to AI / imp 25 · dev 50
- Frontier LLM prices didn't move for 5 months. In August, they moved three times, and one lab tripled its rate. Dev.to AI / imp 55 · dev 40
- Building a Crypto Signal Bot with AI APIs - 2026 Guide Dev.to AI / imp 25 · dev 60
- Groq Free-tier — cheatsheet de limites + failover multi-modelo (PT-BR) Dev.to LLM / imp 20 · dev 50
- Answer Engine Optimization for Insurance Agencies: Six Providers Benchmarked on AI Citations, Lead Quality, and Local Trust Dev.to LLM / imp 25 · dev 40
- Changes to LLM pricing: Baidu, Darkbloom and Tencent Dev.to LLM / imp 35 · dev 35
- Building an AI Brand Monitoring SaaS: How to Track What LLMs Say About Your Brand Dev.to LLM / imp 25 · dev 45
- AI Era Dev.to LLM / imp 20 · dev 70
- Changes to LLM pricing: StreamLake Dev.to LLM / imp 30 · dev 35
- Using LLMs for Crypto Market Analysis in 2026 Dev.to LLM / imp 30 · dev 65
- Run GLM-5.3 Locally: Real Quant Sizes, the llama.cpp Surprise, and the reasoning_effort Trap Dev.to LLM / imp 45 · dev 85
- Your AI Coding Agent Ignores Your Rules File. Here's Which Parts Dev.to LLM / imp 65 · dev 85
- JuryTrace: make agent-judge failures inspectable Dev.to LLM / imp 50 · dev 80
- Changes to LLM pricing: Alibaba, Baidu and StreamLake Dev.to LLM / imp 30 · dev 35
- sub agents being released into my codebase r/ClaudeAI / imp 25 · dev 60
- What's going on with claude usage r/ClaudeAI / imp 30 · dev 50
- A small gift to all my opusfived friends r/ClaudeAI / imp 10 · dev 20
- Fable vs. Astra r/ClaudeAI / imp 40 · dev 70
- I miss when people built things for the internet without turning everything into a subscription r/ClaudeAI / imp 15 · dev 15
- I built a desktop workspace for Claude Code that cuts token usage by 51% r/ClaudeAI / imp 55 · dev 80
- Fable 5.1 one shotted this r/ClaudeAI / imp 20 · dev 40
- I built a Claude Code plugin for my ADHD brain — it remembers the things I forget mid-conversation r/ClaudeAI / imp 50 · dev 75
- GPT vs. Claude: Is the extra intelligence worth the extra cost? r/ClaudeAI / imp 35 · dev 65
- Claude got usage limit reset r/ClaudeAI / imp 25 · dev 50
- What are good examples for artifacts? r/ClaudeAI / imp 20 · dev 60
- Those of you with a finance background, how are creatively using Claude? r/ClaudeAI / imp 20 · dev 40
- We got a limits reset. r/ClaudeAI / imp 20 · dev 45
- I redrew Artificial Analysis using subscription costs instead of API pricing r/ClaudeAI / imp 40 · dev 55
- If you prompt in a language other than English, do you translate your skills? r/ClaudeAI / imp 25 · dev 50
- Control and Answer Claude from your Elgato Stream Deck 🎛️ r/ClaudeAI / imp 30 · dev 60
- Thinking about moving to ChatGPT r/ClaudeAI / imp 25 · dev 60
- Claude Code accidentally nailed a Queen question r/ClaudeAI / imp 15 · dev 30
- Unhinged API Session. Astra vs Fable 5.1 r/ClaudeAI / imp 25 · dev 55
- Why do people still rely on ChatGPT for prompts/instructions for Claude? r/ClaudeAI / imp 25 · dev 45
- What's the next level after the Claude Beginner Stage? r/ClaudeAI / imp 20 · dev 55
- I've designed a set of skills to turn any codebase/repo into a comprehensive coding course! r/ClaudeAI / imp 45 · dev 80
- How can I make my Claude more strategic than tactical? r/ClaudeAI / imp 30 · dev 65
- Is this experience an astral max thing or something? As a plus user on medium and straight vibe coder, I've had to do multiple prompts r/ChatGPTCoding / imp 15 · dev 40
- Best value AI subscription under 20$/200$ (updated for Artificial Analysis Intelligence Index v4.2) r/ChatGPTCoding / imp 35 · dev 55
- What’s the Best Cheap AI Coding Tool? r/ChatGPTCoding / imp 30 · dev 70
- Trying to figure out UIs r/ChatGPTCoding / imp 5 · dev 15
- Benchmarking what agents can do, but what about what agents become? r/ChatGPTCoding / imp 60 · dev 45
- Giving every AI tool the same memory (Claude, ChatGPT, Cursor, Codex, Gemini) r/ChatGPTCoding / imp 55 · dev 70
- title: Everyone on my team mutes the AI code review within two weeks r/ChatGPTCoding / imp 65 · dev 60
- ChatGPT & Codex integration? r/ChatGPTCoding / imp 10 · dev 20
- It must be some kind of psy-op by Anthropic to claim that Fable is anywhere near as good as Astra r/ChatGPTCoding / imp 0 · dev 10
- ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents | US owner of Claude chatbot previously said its models had hacked three organisations during testing r/ChatGPTCoding / imp 70 · dev 55
- Burning Tokens even faster than before r/ChatGPTCoding / imp 15 · dev 10
- how to connect chatgpt + claude + cursor ? r/ChatGPTCoding / imp 8 · dev 25
- Hi peeps, how do you prevent stale test data/code from misleading Claude Code or other coding agents? r/ChatGPTCoding / imp 50 · dev 55
- Editing an old message is the closest thing ChatGPT has to /compact, and it keeps your project chat alive r/ChatGPTCoding / imp 45 · dev 50
- How does ChatGPT + Codex ($20 Plus) hold up for daily custom frontend / WordPress dev before hitting limits? r/ChatGPTCoding / imp 20 · dev 35
- The gap has closed, open source will win r/LocalLLaMA / imp 35 · dev 25
- I've found myself using Local LLM's like 3D printers. r/LocalLLaMA / imp 40 · dev 50
- Qwen3.8-27B beat the Wikipedia game in 6 clicks. r/LocalLLaMA / imp 25 · dev 35
- NInfer vs llama.cpp vs vLLM: quality + speed comparison for Qwen3.8-27B NVFP4 on RTX 5090 r/LocalLLaMA / imp 55 · dev 80
- Qwen3.8 27b for agentic coding and next .... what? r/LocalLLaMA / imp 50 · dev 70
- Otaku — an LLM frontend r/LocalLLaMA / imp 20 · dev 60
- Qwen3.8 Flash Next - Templates Comparison r/LocalLLaMA / imp 35 · dev 75
- You can now run a 90M conversational LLM on the Sony PSP (hardware from 2004). Doesn't get more local than this. r/LocalLLaMA / imp 25 · dev 45
- Any resource on using Blender with local models, and which models work best? r/LocalLLaMA / imp 15 · dev 50
- gfx906-llama-cpp: New PP/TG gains for MI50/MI60/Radeon VII/AMD GCN r/LocalLLaMA / imp 50 · dev 85
- The OpenAI Huggingface incident from an agents POV r/LocalLLaMA / imp 60 · dev 55
- Qwen3.8 27B on Strix - the optimized setup r/LocalLLaMA / imp 45 · dev 80
- LLVM developers begin debate over AGENTS.md for helping AI agents r/LocalLLaMA / imp 75 · dev 85
- RTX 4090 48GB longevity r/LocalLLaMA / imp 15 · dev 30
- I benchmarked 21 Qwen3.8 27B variants on 16GB VRAM r/LocalLLaMA / imp 50 · dev 80
- local vibecoding with Qwen 3.8 27B and Godot r/LocalLLaMA / imp 40 · dev 70
- Is there a local LLM or toolchain to edit 3d models? r/LocalLLaMA / imp 15 · dev 50
- People with "non-enthusaist hardware": how do you use it? r/LocalLLaMA / imp 20 · dev 40
- Qwen3.8-27b is the first Local model im able to blindly trust r/LocalLLaMA / imp 55 · dev 65
- Qwen3.8-Flash-Next (UD-Q4_K_XL) on a single RTX 3090 24GB + 128GB DDR4, is this config optimal? r/LocalLLaMA / imp 40 · dev 80
- Can humans approve AI actions mid-call? r/AI_Agents / imp 55 · dev 50
- I read 40k words of AI output a day. Here's how to stop reading most of it. r/AI_Agents / imp 65 · dev 60
- Nothing is a black box, agent transparency matters more than capabilities right now. r/AI_Agents / imp 65 · dev 55
- Cekura / Cyara / TestMu Agent Testing, are these even solving the same problem? r/AI_Agents / imp 55 · dev 65
- Built a fully autonomous Android agent — 30 tools, on-device model, and a policy gate that blocks every untraced action r/AI_Agents / imp 75 · dev 80
- From Gemini->ChatGPT r/AI_Agents / imp 15 · dev 25
- messa is live, agent over text r/AI_Agents / imp 35 · dev 55
- Cheap subscription for coding microservices r/AI_Agents / imp 12 · dev 30
- Should an AI router be allowed to make the final execution decision? r/AI_Agents / imp 60 · dev 65
- The bottleneck in agent-run client work isn't building the automation, it's the exception queue r/AI_Agents / imp 65 · dev 55
- AI for a podcast r/AI_Agents / imp 10 · dev 35
- What if your AI agent could spend 50% less on tokens — and actually recover when things go wrong? r/AI_Agents / imp 65 · dev 70
- Astra sucks or skill issue? - did not follow orders like Sol does r/AI_Agents / imp 0 · dev 20
- How are people actually building and selling AI agents? What should a TS/full-stack dev learn? r/AI_Agents / imp 50 · dev 65
- Am I the only one thinking AI workflows are more of a burden rather than a relief? r/AI_Agents / imp 55 · dev 55
- Looking for advice to automate Tamil Q&A formatting in MS Word (Phonetic Font & Options Generation) r/AI_Agents / imp 10 · dev 40
- Apparently those agents can sometimes quietly return wrong numbers for over a week before anyone notices. Seems like a common enough issue to watch out for. r/AI_Agents / imp 70 · dev 60
- folks, gonna be launching on #producthunt in a weeks time. Any tips on how to get a successful launch done? r/AI_Agents / imp 15 · dev 20
- Does A2A actually make agents interoperable? r/AI_Agents / imp 65 · dev 75
- What is inside the Ahrefs AI tools now, for anyone who last looked a year ago r/AI_Agents / imp 30 · dev 50
- Shipping agents that act on customers' real social accounts: the guardrails we ended up needing r/AI_Agents / imp 75 · dev 70
- Best Tools For Building AI Agent? r/AI_Agents / imp 15 · dev 50
- AI Vendors Don’t Lock You In. They Lock You Out. r/AI_Agents / imp 75 · dev 65
- sub agents being released into my codebase r/OpenAI / imp 55 · dev 70
- Fable 5.1 vs GPT 6 Astra, 3D Blender, mind blowing difference! r/OpenAI / imp 0 · dev 30
- Differences Between GPT-5.6 Sol Pro and GPT-6 Astra Pro on MineBench.ai r/OpenAI / imp 0 · dev 25
- Astra Minecraft houses at different reasoning levels r/OpenAI / imp 0 · dev 20
- Astra (GPT-6) High Intelligence, Low Intuition r/OpenAI / imp 0 · dev 15
- Made in 2h with astra :o r/OpenAI / imp 10 · dev 25
- P(DOOM): AI safety parody of DOOM 64 r/OpenAI / imp 10 · dev 20
- 3D Scenes in 20 minutes r/OpenAI / imp 10 · dev 25
- My first "holy shit" moment with GPT-6 Astra: I asked it to create a world in Unreal Engine, and fill it with humans (each an Astra-powered agent) who all have to work together to survive. r/OpenAI / imp 15 · dev 30
- thanks to astra i finally gathered all the infinity stones r/OpenAI / imp 5 · dev 15
- Astra can play games! r/OpenAI / imp 15 · dev 40
- Nobody is Talking About GPT 6 Astras Massive Hallucination Improvements r/OpenAI / imp 40 · dev 35
- Be Aware: Astra Burns Usage r/OpenAI / imp 50 · dev 45
- This is gold. Tibo is the GOAT. r/OpenAI / imp 5 · dev 10
- Even ASTRA cannot do drag&drop in our system r/OpenAI / imp 25 · dev 40
- ChatGPT Ads rolling out to addition countries including the MENA region r/OpenAI / imp 35 · dev 30
- Why are you guys using OpenAI over Anthropic? r/OpenAI / imp 20 · dev 25
- Is this enough to have us worried about world changing consequences for jobs? Or do we need ASI for that? r/OpenAI / imp 40 · dev 20
- Language Models Can Control Their Own Attention [R] r/MachineLearning / imp 60 · dev 80
- NeurIPS 2026 Automatic Reference Checker [R] r/MachineLearning / imp 15 · dev 40
- What is the general design of these new math solving systems? [D] r/MachineLearning / imp 55 · dev 75
- Implementing Embedding Gemma from scratch in PyTorch [P] r/MachineLearning / imp 50 · dev 85
- Gpt 5,6,7: Does it even matter? The (ghost) productivity question. [D] r/MachineLearning / imp 55 · dev 55
- OpenAI CEO Sam Altman says 38,000 ChatGPT queries use as much water as the production of one almond — says data centers use no more water than an office building: “For every 38,000 ChatGPT queries, that is the same amount of water that is used in the production of single almond in California.” r/artificial / imp 40 · dev 20
- Why are more people not concerned about privacy? r/artificial / imp 35 · dev 15
- OK, who's gonna make a model of Michael Knight's computer KITT ? r/artificial / imp 5 · dev 15
- Am I the only one thinking AI workflows are more of a burden than relief? r/artificial / imp 55 · dev 55
- The generation generation. r/artificial / imp 30 · dev 10
- What if AI models are designed to make mistakes? 🤔 r/artificial / imp 25 · dev 20
- I'm building a 2D video to 3D animation translation layer. r/artificial / imp 40 · dev 65
- What's the deal with all those cheap 3D modelled web apps? r/artificial / imp 20 · dev 30
- Bernie Sanders Wants to 'Pause AI Development NOW': Dwarkesh Patel Asks 'Pause to Do What?' r/artificial / imp 25 · dev 10
- [Discussion] Do you think AI can develop secure enough projects? r/artificial / imp 55 · dev 60
- Top AI Development Companies to Consider in 2026 r/artificial / imp 15 · dev 35
- Study: Generative AI is making writing on Reddit and elsewhere boring r/artificial / imp 50 · dev 40
- [Experiment] I trained a model on childhood photos to simulate memory recall r/artificial / imp 20 · dev 50
- Are we finally hitting the “Scaling-Wall”? Test-Time Compute meta on AI Industry. r/artificial / imp 65 · dev 70
- Salesforce blames its Claude addiction for denting profit margin guidance r/artificial / imp 55 · dev 30
- How to gamify jobs which are inherently not fun r/artificial / imp 15 · dev 20
- Can we all acknoledge how crazy AI is? r/artificial / imp 20 · dev 10
- 50.5% of Americans Say AI Romance Can Count as Cheating r/artificial / imp 15 · dev 5
- AI agents can now pay for things online by themselves. Here is what can go wrong, and what I built to catch it r/artificial / imp 75 · dev 70
- Why would you use an AI to write your replies on Reddit? r/artificial / imp 15 · dev 10
- ChatGPT Plus vs Gemini AI Pro vs Claude Pro for serious academic research r/artificial / imp 25 · dev 40
- Automatic AI Responses to Customer Service Issues r/artificial / imp 20 · dev 15
- Child sexual abuse survivor alleges Elon Musk’s AI chatbot used photos of her to generate new illegal images r/artificial / imp 40 · dev 20
- Using Blender with coding agents on macOS Simon Willison / imp 35 · dev 75
- The Pelican comparison grid for Astra is pretty interesting Simon Willison / imp 25 · dev 30
- August newsletter is out Simon Willison / imp 20 · dev 15
- OpenClaw Power, MacBook Simplicity: Five Days With Grok Bot Latent Space / imp 40 · dev 50
- [AINews] GPT-6 Astra: OpenAI’s biggest LLM launch of all time Latent Space / imp 75 · dev 80
- Experiential Labs Product Hunt / imp 30 · dev 65
- Ponytail Product Hunt / imp 25 · dev 70
- GitWarren Product Hunt / imp 30 · dev 75
- Show HN: MobileCode – OpenCode with Built-In iOS and Android Previews HN DeepSeek/Grok/OpenCode / imp 35 · dev 80
- Eidon: Self-hosted AI assistant with Grok Bot-like agent mode HN DeepSeek/Grok/OpenCode / imp 25 · dev 70
- Designing Grok Bot for a world of persistent agents HN DeepSeek/Grok/OpenCode / imp 65 · dev 85
- This Week's Top Five Stories in AI - AI Magazine Google News DeepSeek / imp 15 · dev 20
- DeepSeek: How Has It Disrupted Global Open Source AI - AI Magazine Google News DeepSeek / imp 50 · dev 60
- Corporate America Is Swapping Claude and GPT for Cheaper DeepSeek Models - Startup Fortune Google News DeepSeek / imp 45 · dev 30
- Meta’s Flagship AI Model Makes Stunning Comeback: Cheaper Than DeepSeek, Chinese-American Top Executive Challenges Gemini Head-On - 36 Kr Google News DeepSeek / imp 50 · dev 50
- Why million token context AI agents matter - StartupHub.ai Google News DeepSeek / imp 60 · dev 80
- How to Use Qwen AI: 13 Steps, 90 Min [2026] - tech-insider.org Google News DeepSeek / imp 20 · dev 60
- DeepSeek Plans a Massive Huawei AI Chip Deployment - Techgenyz Google News DeepSeek / imp 45 · dev 40
- Cracking 1.33 Trillion Daily Tokens: B.AI Powers the "AI Grid" with Full-Stack Infrastructure to Fuel the Agentic Era - GlobeNewswire Google News DeepSeek / imp 65 · dev 75
- Open Models Now Handle 80-90% of Enterprise AI Tokens, Ollama's Jeffrey Morgan Says - finance.biggo.com Google News DeepSeek / imp 50 · dev 70
- Setting Grok Bot loose on procurement - X.ai Google News Grok/xAI / imp 45 · dev 50
- Judge rejects Musk bid to halt Minnesota ban on AI nudifying - Courthouse News Google News Grok/xAI / imp 40 · dev 20
- This Grok portfolio just destroyed the S&P 500 - Finbold Google News Grok/xAI / imp 20 · dev 35
- Starlink is offering half-price internet near xAI data centers, but critics want the pollution fixed - TechSpot Google News Grok/xAI / imp 15 · dev 15
- Tesla Optimus, Grok, Cybercab: Cutting Through the TSLA Hype - MarketWise Google News Grok/xAI / imp 25 · dev 30
- GPT-6 Astra vs Grok 4.6 vs Qwen3.8-Max: 8x Price Gap [2026] - tech-insider.org Google News Grok/xAI / imp 55 · dev 75
- Grok Bot Template Marketplace: Everything You Need to Know - BASENOR - Tesla Accessories Google News Grok/xAI / imp 35 · dev 70
- Victims And Cities Battle Flood Of AI Child Sexual Abuse Deepfakes Involving Pictures Of Real People - Patch Google News Grok/xAI / imp 50 · dev 25
- AIが問いを立て直した ── 音声のズレを追って感じたLLMの進化 Zenn LLM / imp 0 · dev 80 (いいね相当スコア: 0)
- 要件定義書をLLMと一緒に書いてみて、うまくいったこと・ダメだったこと Zenn LLM / imp 0 · dev 75 (いいね相当スコア: 1)
- AI向けの目次を置いて14日ログを取ったら、148回取られてトップ以外は誰も読まなかった Zenn LLM / imp 0 · dev 65 (いいね相当スコア: 0)
- Claude Code で一番高いリクエストは「さっきの続きやって」かもしれない Zenn LLM / imp 0 · dev 75 (いいね相当スコア: 0)
- OpenAI API 入門 — API キーの取得から初回リクエストまで Zenn LLM / imp 0 · dev 70 (いいね相当スコア: 0)
- GPT-6 Astra は再現性の高いコードを書く——Sol との比較 240 試行で見えた事実(オトナの自由研究 #38) Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 0)
- 【今さら聞けない】MCPを周回遅れで学び直す〜AWSでの活用とAIエージェントの共通規格〜 Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 0)
- MCP防御で本番事故を防ぐためのチェックリスト Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 1)
- 生成AIに送ったデータはどこへ行く?Amazon Bedrockのセキュリティを整理する Zenn LLM / imp 0 · dev 80 (いいね相当スコア: 2)
- 「スコア上位から選ぶ」だけでは多様性を担保できない ― LLMランキングに業種ラウンドロビンを入れた話 Zenn LLM / imp 0 · dev 65 (いいね相当スコア: 2)
- ファインチューニングとRAGの使い分け:どちらで解くべき問題かを初心者向けに整理する Zenn LLM / imp 0 · dev 80 (いいね相当スコア: 0)
- 誰もやってないベンチマークとAPI非依存におけるアプリ内アストラの挙動報告 Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 1)
- 通常のファインチューニングとLoRAでGPUメモリ消費を比較検証 Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 0)
- LLMを呼ばないための実装に時間を使ったら、変動費が9割減った Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 0)
- 【Google】Gemini 3.8 Flash & Cyber登場!特徴とAPI活用法 Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 1)
- AI彼女アプリを作っていて気付いた。気持ちと約束と記憶は、同じ速さでは残らなかった Zenn LLM / imp 0 · dev 70 (いいね相当スコア: 0)
- Qwen3.8-Flash-Next量子化版は最小72.5GB、同じ重みで出力も5割変わる Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 0)
- なぜフロンティアモデルのクラウドLLMは弱くなるのか ─ 『最初はよかったのに』を分解する Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 1)
- LangChain.js × Gemini × MCP の Schema Error を回避する方法(@langchain/google版) Zenn LLM / imp 0 · dev 85 (いいね相当スコア: 0)
- 【備忘録】初めてスマホアプリを構築・ビルドしてみた Zenn LLM / imp 0 · dev 75 (いいね相当スコア: 2)
- 【音響物理×LLM】耳介の干渉ノッチをRoPEに拡張し、類似トークンを識別・消去する「SN-RoPE」 Zenn 機械学習 / imp 0 · dev 90 (いいね相当スコア: 1)
- 実写画像のOCRで、CPU専用モデルを2つに分けた理由 Zenn 機械学習 / imp 0 · dev 85 (いいね相当スコア: 0)
- TimesFM 3.0 でスーパーの米価格を予測する Zenn 機械学習 / imp 0 · dev 75 (いいね相当スコア: 0)
- ローカル環境でWhisperXを使い、2人の会話を話者別に議事録化する Zenn 機械学習 / imp 0 · dev 80 (いいね相当スコア: 3)
- Reservoir Computing の基本と簡易検証 Zenn 機械学習 / imp 0 · dev 80 (いいね相当スコア: 1)
- Claude Codeと作る競馬予測システム開発記(4) 買い目戦略と評価編 Zenn 機械学習 / imp 0 · dev 80 (いいね相当スコア: 0)
- AIにデータを集めさせる前に、3つの検査を決める Zenn 機械学習 / imp 0 · dev 80 (いいね相当スコア: 0)
- BigQuery MLの需要予測モデルでEC仕入れ量を最適化する実装手順 Zenn 機械学習 / imp 0 · dev 85 (いいね相当スコア: 0)
- MobileNetV2 を手書き NEON で速くする — 遅い層を個別に潰して ORT を抜く Zenn 機械学習 / imp 0 · dev 90 (いいね相当スコア: 0)
- AIに競艇予想モデルを作らせる。その数字を信じるのは人間の仕事です Zenn 機械学習 / imp 0 · dev 80 (いいね相当スコア: 0)
- 【LLM】ベンチマーク用オリジナル上級問題(v1) Qiita LLM / imp 0 · dev 80 (いいね相当スコア: 0)
- 同じ指示、違う画像。画像生成AIモデルの「解釈差」を観測・比較してみた Qiita LLM / imp 0 · dev 75 (いいね相当スコア: 0)
- vLLM + LiteLLM を本線に、セルフホストAIワークスペースをワンコマンドで立てる手順(Ollama は最短ルート) Qiita LLM / imp 0 · dev 85 (いいね相当スコア: 0)
- AIエージェントの評価を実装する:Evalsの3層構造とpytest設計 Qiita LLM / imp 0 · dev 90 (いいね相当スコア: 0)
- CLIPのゼロショットでは当たらない「自分の分類軸」を、線形プローブで作る Qiita 機械学習 / imp 0 · dev 85 (いいね相当スコア: 0)
- OpenAIが「GPT-6 Astra」を発表、コーディング・サイバーセキュリティで最先端性能を主張 Qiita 機械学習 / imp 0 · dev 80 (いいね相当スコア: 0)
- 【AIの取扱説明書 第1章】AIは流暢に間違える ― 生成AIと付き合うために知っておきたい3つの特性 Qiita 機械学習 / imp 0 · dev 60 (いいね相当スコア: 0)
- TensorFlowで作ったモデルをVertex AI(現Agent Platform)にデプロイして推論を試してみた Qiita 機械学習 / imp 0 · dev 85 (いいね相当スコア: 0)
- AIのいいなりになってAIを学ぶ Step5 Gate02:Agentに必要なPythonだけを覚える note LLM / imp 0 · dev 70 (いいね相当スコア: 取得失敗)
- 第四話「三巫女に聞いてみよう」 note LLM / imp 0 · dev 40 (いいね相当スコア: 取得失敗)
- データ価値のパラダイムシフトと知識蒸留:情報熱力学および論理的検証器を通じた二次データ生成のメカニズムに関する包括的考察 note LLM / imp 0 · dev 75 (いいね相当スコア: 取得失敗)
- 似ているようで似ていない高度化と高次化が示すAIの未来 note LLM / imp 0 · dev 60 (いいね相当スコア: 取得失敗)
- ClaudeとGoogle Workspaceをどう連携する?公式コネクタ・MCP・gws CLIの使い分けを整理できる実務手引き note LLM / imp 0 · dev 85 (いいね相当スコア: 取得失敗)
- medgap-qwen-3.5-9b Phase 5 インジェクション後編レポート note LLM / imp 0 · dev 80 (いいね相当スコア: 取得失敗)
- medgap-qwen-3.5-9b Phase 4 インジェクション前編レポート note LLM / imp 0 · dev 80 (いいね相当スコア: 取得失敗)
- medgap-qwen-3.5-9b 総合ベンチマークレポート(全52問) note LLM / imp 0 · dev 80 (いいね相当スコア: 取得失敗)
- medgap-qwen-3.5-9b Phase 2 日本語性能レポート note LLM / imp 0 · dev 80 (いいね相当スコア: 取得失敗)
- AIにとって言葉が「自分事」になるとは何か――Synthetic Linguistic Agencyという試み note LLM / imp 0 · dev 70 (いいね相当スコア: 取得失敗)
- GPT-6 Astraで見えたAIの設計図――知能を測る単位はLLMからシステム全体へ note LLM / imp 0 · dev 90 (いいね相当スコア: 取得失敗)
- 生成AIの収支崩壊:2027年問題とトークン経済の真実 Ed Zitron note LLM / imp 0 · dev 40 (いいね相当スコア: 取得失敗)
- 【生成AIニュース+】『LLaDA-Image ComfyUI』『LLaDA-Image-Turbo ComfyUI』『K2-Horizon-375B-A23B』『Tripo P2.0 Preview マルチビュー』『DLSS 5 Visual Enhancer』『ComfyUI NVIDIA DLSS Frame Interpolation』『Layers Studio』『Camera to Blender』『MiniMax H3 VR180 SBS LoRA』『Viggle-Animate』他 note LLM / imp 0 · dev 30 (いいね相当スコア: 取得失敗)
- 「AI夜市」参加レポート:カオスな現場で知った、AI技術の“裏側”の尖り方と、技術者の情熱の原点 note LLM / imp 0 · dev 50 (いいね相当スコア: 取得失敗)
- 【第1回】テキストからビジュアルへ、そして再びテキストへーー生成AIが変えるコンピュータとの付き合い方 note LLM / imp 0 · dev 60 (いいね相当スコア: 取得失敗)
- 宇宙で野菜を育てる話→無人の太陽系工業圏まで考えてしまった note LLM / imp 0 · dev 0 (いいね相当スコア: 取得失敗)
- AIコーディングのゲームデモをどう読むか note LLM / imp 0 · dev 85 (いいね相当スコア: 取得失敗)
- GPT6はAGIなのか?【レポート11】 note LLM / imp 0 · dev 75 (いいね相当スコア: 取得失敗)
- 推論モデルを「迷子」にさせない思考トポロジー設計——多段階推論(Reasoning)の暴走を抑え込む3つのアンチパターン脱却術 note LLM / imp 0 · dev 90 (いいね相当スコア: 取得失敗)
- 「確率の半歩先」読了 note LLM / imp 0 · dev 15 (いいね相当スコア: 取得失敗)
- 情報機関の方法論と市場調査:OSINT・HUMINT・SIGINT・MASINT・FININT の企業応用と倫理的境界 note LLM / imp 0 · dev 50 (いいね相当スコア: 取得失敗)
- DAY48|AIに任せるのではない。AIと一緒に判断する|Company AI OS開発記録 note LLM / imp 0 · dev 85 (いいね相当スコア: 取得失敗)
- プロンプト設計の記事をAIに書かせたら事実検証ゲートで4件引っかかった note LLM / imp 0 · dev 80 (いいね相当スコア: 取得失敗)
- 😈4ヶ月ハルシネーションを続けたGeminiとの直近対話を✴️サカナAIフグで解析版‼️ note LLM / imp 0 · dev 70 (いいね相当スコア: 取得失敗)