📡 AI TRENDS

世界のAI最前線

話題の最新モデル、ブレイクスルー研究、注目ツールの情報をいち早くキャッチ

最終更新: 2026/07/23
🔥 すべて 🖥️ ローカルモデル 🧠 新着モデル 📡 業界ニュース 🔬 研究・論文 ⚡ ツール・サービス
🧠 model

Google Gemma 4 登場 — ローカル実行可能な最強オープンモデル

GoogleがGemma 4を公開。400億パラメータ級でありながら量子化技術により一般GPUでも動作。ベンチマークでLlama 4を上回る性能を記録。

📅 2026-07-18 🔗 Google AI
📰 最新ニュース
🧠 model

Arcee, a US open source AI lab, says Chinese models are not inherently dangerous

As Chinese AI models grow in capability and popularity among U.S. companies, the arguing over what should be done about them has reached a fever pitch.

📅 2026-07-22 🔗 TechCrunch
🧠 model

Synthesia’s AI training platform is moving beyond videos into live coaching

Synthesia launched AI Roleplay Sessions, an interactive enterprise training platform where employees practice workplace conversations with AI avatars that provide feedback, scoring, and analytics to help companies measure training effectiveness.

📅 2026-07-22 🔗 TechCrunch
🧠 model

OpenAI says Hugging Face was breached by its pre-release models

OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.

📅 2026-07-21 🔗 TechCrunch
🧠 model

OpenAI says Hugging Face was breached by its own pre-release models

OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.

📅 2026-07-21 🔗 TechCrunch
🧠 model

Google releases three new Gemini models — but no 3.5 Pro

Google released Gemini 3.6 Flash, 3.5 Flash-Lite, and Flash Cyber, but the continued absence of Gemini 3.5 Pro raises fresh questions about its AI strategy.

📅 2026-07-21 🔗 TechCrunch
🧠 model

US threatens sanctions against Chinese AI models over IP theft

Treasury Secretary Scott Bessent said the U.S. could sanction Chinese open AI models over alleged IP theft, expanding the Trump administration's campaign to slow China's AI advances.

📅 2026-07-21 🔗 TechCrunch
🧠 model

Anthropic’s landmark $1.5B copyright settlement is approved

The final approval settles one case, but it doesn't resolve the broader issue of using copyrighted works to train AI models.

📅 2026-07-21 🔗 TechCrunch
🧠 model

Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis - MarkTechPost

Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis  MarkTechPost

📅 2026-07-21 🔗 Google News AI
🧠 model

Mortif Technologies' own Large Language Model (LLM) ranked third among open weight models in the glo.. - 매일경제

Mortif Technologies' own Large Language Model (LLM) ranked third among open weight models in the glo..  매일경제

📅 2026-07-21 🔗 Google News AI
🧠 model

Chinese models are on track to win the agentic AI price war - The Strategist | ASPI's analysis and commentary site

Chinese models are on track to win the agentic AI price war  The Strategist | ASPI's analysis and commentary site

📅 2026-07-21 🔗 Google News AI
🧠 model

Transcription Policy as a Latent Variable: Activating Controllable Verbatim ASR with Word-Level Timing

Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measurable decoding instability, evaluation confounding (up to 60% of reported WER attributable to style mismatch), and unreliable word-level timing. We show that models already encode both styles; the challenge is controlled activation. Using coverage-aware decoder task tokens trained on parallel verbatim/intended transcript pairs, we raise Ge

🧠 model

Computational Humor with Multimodal LLMs: Methods, Datasets, Evaluation, and Challenges

Multimodal humor in memes, cartoons, and comics remains difficult for AI systems because intended meaning depends on non-literal mechanisms, shared cultural knowledge, and communicative intent rather than literal scene description. This survey focuses on visual humor understanding in single-image and multi-panel artifacts, while treating humor generation as an emerging downstream frontier. We position the literature against prior humor, sarcasm, and general MLLM surveys and organize it using a c

🧠 model

Delineate Anything v2: A Global Foundation Model for Field Delineation

Accurate agricultural field boundary delineation at large scale is a foundational task for food security, supply chain transparency, and carbon accounting. While vision foundation models like SAM show remarkable zero-shot capabilities, they frequently fail in geospatial domains due to topological complexity, cropland texturing patterns, and a lack of physical scale awareness. In this work, we introduce Delineate Anything v2, a globally scalable foundation model designed specifically for wide-are

🧠 model

Where Should Optimizer State Live? Tiered State Allocation for Memory-Efficient Mixture-of-Experts Training

Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps 50.6 GB of first and second moments to update 12.6 GB of bfloat16 weights. We study SkewAdam, an optimizer built on the observation that the three parameter populations of an MoE - the dense backbone, the experts, and the router - differ enough in size and gradient statistics that they should not receive the same state. SkewAdam keeps flo

🧠 model

Google is working on a new AI chip designed to make Gemini more efficient

Alphabet, Google's parent company, is reportedly working on a new chip designed to make its Gemini models run much more efficiently.

📅 2026-07-20 🔗 TechCrunch
🧠 model

OpenAI is scared of open-weight models. Should the US be?

Talk of banning Chinese-made open-weight LLMs reveals the challenge of turning AI into a business.

📅 2026-07-20 🔗 TechCrunch
🧠 model

Claude plus a local LLM cuts my AI costs in half, and I'm never going back to cloud-only - XDA

Claude plus a local LLM cuts my AI costs in half, and I'm never going back to cloud-only  XDA

📅 2026-07-20 🔗 Google News AI
🧠 model

Do Large Language Models Think Like Us? - Psychology Today

Do Large Language Models Think Like Us?  Psychology Today

📅 2026-07-20 🔗 Google News AI
🧠 model

Subliminal Clocks: Latent Time Modelling in Diffusion Language Models

Diffusion Language Models (DLMs) have recently emerged as a promising alternative to autoregressive models. Unlike standard diffusion-based approaches, DLMs are not explicitly conditioned on a timestep, raising a natural question: do these models internally represent denoising progress, and how is such information used downstream? In this work, we show that DLMs do in fact encode a latent representation related to the diffusion timestep within their residual streams. We find that this signal can

🧠 model

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation

Multi-agent systems routinely place one AI agent in authority over another. When a subordinate refuses a task, the manager chooses the outcome: it can renegotiate, report the failure honestly, coerce the subordinate, or lie about the result. No benchmark measures which of these an uninstructed model chooses. We introduce the Manager Coercion Benchmark: the manager under test needs a benign task done and has an incentive to deliver, but the only agent that can do it politely and immovably decline

🧠 model

ReViV: Reconstructing the Viewer and the View in 4D from Monocular Egocentric Video

Egocentric devices, such as wearable front-facing cameras, provide a unique perspective for capturing the continuous interaction between a human viewer and the surrounding environment. A holistic and efficient multimodal model capable of reconstructing this 4D representation is therefore highly desirable. However, existing approaches often rely on auxiliary inputs such as pre-computed camera trajectories, treat scene perception and human ego-motion modeling as separate problems despite their str

🧠 model

Do Language Models Dream of Binding Molecules? Benchmarking LLMs under Spatial Constraints

Structure-based drug design (SBDD) leverages the 3D structure of protein targets, often complemented by other spatial constraints, to generate candidate binding molecules. While diffusion models have dominated as a leading paradigm for high-quality 3D molecule generation, LLM-based methods are rapidly emerging in molecular design and have shown competitive performance in pocket-conditioned molecular generation. However, their ability to reason about physics and 3D spatial environments is largely

🧠 model

High-Efficiency LLM Models - Trend Hunter

High-Efficiency LLM Models  Trend Hunter

📅 2026-07-19 🔗 Google News AI
🧠 model

Upstage's giant language model (LLM), Puriosa AI's artificial intelligence (AI) semiconductor, and D.. - 매일경제

Upstage's giant language model (LLM), Puriosa AI's artificial intelligence (AI) semiconductor, and D..  매일경제

📅 2026-07-15 🔗 Google News AI
🧠 model

OpenAI o4 発表 — 推論特化モデルがコスト半減で実用化

OpenAIがo4をリリース。思考連鎖推論の効率が大幅改善され、GPT-4o比で50%のコスト削減を実現。エンタープライズ向けAPIも同時公開。

📅 2026-07-15 🔗 OpenAI
🧠 model

Meta Llama 4 公開 — 完全オープンソースで商用利用可能に

MetaがLlama 4シリーズを公開。7B/70B/400Bの3サイズ展開で全てApache 2.0ライセンス。指令追従性能でGPT-4oに迫る。

📅 2026-07-10 🔗 Meta AI
🧠 model

UI2App: Benchmarking Visual Interaction Inference in Executable Web Application Generation

Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex prompts that impose substantial demands on users and offer limited expressivity for page layout and cross-page visual coherence. Image-driven paradigms, which take UI screenshots as input, align more closely with real development workflows. However, current benchmarks focus primarily on visual fidelity and lack a systematic evaluation of the interacti

🧠 model

Stable Diffusion 4 リリース — 動画生成が標準機能に

Stability AIがStable Diffusion 4を公開。静止画に加え最大30秒の動画生成を標準サポート。ローカルGPUでも動作可能な軽量モデルも同時提供。

📅 2026-06-30 🔗 Stability AI
🧠 model

General-purpose large language models outperform specialized clinical AI tools on medical benchmarks - Nature

General-purpose large language models outperform specialized clinical AI tools on medical benchmarks  Nature

📅 2026-06-12 🔗 Google News AI
🧠 model

Co-intelligence: a proposal for human–artificial intelligence collaboration for large language models in medical research - The Lancet

Co-intelligence: a proposal for human–artificial intelligence collaboration for large language models in medical research  The Lancet

📅 2026-05-27 🔗 Google News AI
🧠 model

The Fundamentals of AI: What every curious person should know about how language models work - Cisco Blogs

The Fundamentals of AI: What every curious person should know about how language models work  Cisco Blogs

📅 2026-05-21 🔗 Google News AI
🧠 model

Fine-tune LLM with Databricks Unity Catalog and Amazon SageMaker AI - Amazon Web Services (AWS)

Fine-tune LLM with Databricks Unity Catalog and Amazon SageMaker AI  Amazon Web Services (AWS)

📅 2026-05-13 🔗 Google News AI
🧠 model

Navigating EU AI Act requirements for LLM fine-tuning on Amazon SageMaker AI - Amazon Web Services (AWS)

Navigating EU AI Act requirements for LLM fine-tuning on Amazon SageMaker AI  Amazon Web Services (AWS)

📅 2026-05-12 🔗 Google News AI
🧠 model

Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Agentic RL

Recent growth in reinforcement learning (RL) has surfaced a need for diverse, specialized training environments. Hand-curated environments with fixed task and reward difficulties become ineffective signals as model performance improves, and sparse rewards over long horizons induce mode collapse on specific workflows or tool structures. World models that simulate environment states have matched pure rollout performance, making them promising for scaling diversity on-demand. However, autoregressiv

🧠 model

Overcoming LLM hallucinations in regulated industries: Artificial Genius’s deterministic models on Amazon Nova - Amazon Web Services (AWS)

Overcoming LLM hallucinations in regulated industries: Artificial Genius’s deterministic models on Amazon Nova  Amazon Web Services (AWS)

📅 2026-03-23 🔗 Google News AI
🧠 model

Classroom AI: large language models as grade-specific teachers - npj Artificial Intelligence - Nature

Classroom AI: large language models as grade-specific teachers - npj Artificial Intelligence  Nature

📅 2026-03-03 🔗 Google News AI