Filter

×
Active Filters Clear All
Keyword: Token ×
177 Total Reports
9/9 Page
OpenAI Other 2026-03-05

OpenAI Launches GPT-5.4 Model Focused on Professional Work Scenarios

OpenAI releases GPT-5.4, designed for professional work with enhanced code generation and tool search. Key upgrade includes 1M token context window for long document and complex code processing. Positioned as an efficient assistant without architectural changes.

Google Other Medium Signal 2026-03-04

Google Launches Efficient Inference Model Gemini 3.1 Flash-Lite

Google released Gemini 3.1 Flash-Lite, optimized for high-frequency workloads with 2.5x faster first-token response and 45% higher output speed. Available via AI Studio and Vertex AI, it features thinking depth adjustment for scalable AI applications like translation and content moderation.

Trend Micro Other High Signal 2026-03-03

Trend Micro Report Highlights AI Supply Chain Risks and Model Attack Surfaces

Trend Micro's 'Fault Lines in the AI Ecosystem' report systematically analyzes security risks in the AI supply chain, including training data poisoning, third-party plugin vulnerabilities, and model theft attacks. It indicates that enterprise AI security boundaries have expanded from traditional IT infrastructure to the model layer and data pipelines.

Fortinet Product Launch High Signal 2026-03-01

FortiOS 8.0 FortiAI: Deep Dive into RAG-Powered Intelligent O&M Assistant

FortiOS 8.0 introduces FortiAI-Assist, a RAG-based AI assistant embedded in FortiOS, providing documentation Q&A, troubleshooting, and CLI command generation. Supports dual AI providers with token-based billing.

Cisco Other High Signal 2026-02-10

Cisco Launches G300 Chip and Systems for AI Agent-Era Data Center Networking

Cisco introduces 102.4Tbps Silicon One G300 switching chip with liquid-cooled N9000/8000 systems delivering 70% energy efficiency, 1.6T optics support, and Nexus One unified management plane upgrade.

Trend Micro Other High Signal 2026-01-07

Trend Micro Reveals Novel Docker Desktop WSL2 VM Escape Attack Surface

Trend Micro has discovered novel virtual machine escape techniques in Docker Desktop under WSL2, allowing attackers to leverage exposed internal APIs and configuration mechanisms to break out of the container environment and execute arbitrary code on the host. This exposes serious security boundary risks hidden within development toolchains.

NVIDIA Other 2025-11-08

NVIDIA Launches Interactive AI Agent for GPU-Accelerated Data Science with Nemotron Nano-9B

NVIDIA unveils an interactive AI agent powered by Nemotron Nano-9B-v2 and CUDA-X libraries, enabling natural language orchestration of ML workflows. It achieves 3x-43x GPU acceleration over CPU for data processing, model training, and hyperparameter optimization.

NVIDIA Other 2025-06-06

NVIDIA and SK hynix Co-Architect Next-Gen Memory for AI Factories, Locking HBM4 to Vera Rubin

NVIDIA and SK hynix announce a multi-year tech partnership to co-develop next-gen memory for Vera Rubin, RTX Spark, and Jetson Thor. Separately, SK Telecom deploys a gigawatt-scale AI cloud using the full DGX stack, targeting 2027. This elevates SK hynix from supplier to co-architect, strengthening NVIDIA's lock-in on HBM and the AI ecosystem.

NVIDIA Other 2025-06-01

NVIDIA RTX Spark and Nemotron-3 Ultra: AI Control Shifts from Cloud to Personal Edge

NVIDIA launched RTX Spark personal AI supercomputer (co-developed with MediaTek) and Nemotron-3 Ultra open-source model at GTC Taipei 2026. The N1X chip delivers 1 PFLOPS local AI compute, bringing LLM inference to PCs. This marks NVIDIA's pivot from cloud GPU vendor to edge AI infrastructure monopolist, redefining the PC as an AI-native device.

Microsoft Other Medium Signal 2025-02-27

Microsoft Launches Phi-4 SLM Series to Enhance Edge AI and Multimodal Reasoning

Microsoft introduced the Phi-4 family of small language models (SLMs), featuring the 5.6B-parameter Phi-4-multimodal capable of processing speech, vision and text. The models are now available in Azure AI Foundry, HuggingFace and NVIDIA's API Catalog with optimized edge computing capabilities.

Anthropic Other 2021-10-07

Anthropic Launches Project Glasswing: AI Model Autonomously Finds Zero-Days, Reshaping Cyber Defense

Anthropic announces Project Glasswing, partnering with AWS, Apple, Cisco, Google, Microsoft, NVIDIA, and others to use its frontier model Claude Mythos Preview for autonomous vulnerability discovery. The model found thousands of zero-days, including decades-old flaws in OpenBSD, FFmpeg, and Linux kernel. Anthropic commits $100M in usage credits, aiming to shift cybersecurity to AI-driven defense at scale.

Google Other High Signal 2020-10-11

Google Cloud Integrates MCP with Apigee and Advances Agentic Platform to Evolve Enterprise APIs for AI Agents

Google Cloud announced the general availability of Model Context Protocol (MCP) in Apigee and the advancement of its Agentic Platform, aiming to transform traditional enterprise APIs into secure, governed tools for AI agents at scale. This move integrates API governance, security layers, and AI inference infrastructure, providing core platform capabilities for enterprises shifting from API-driven to agent-driven architectures.

Trend Micro Other High Signal 2020-06-01

Trend Micro Exposes Azure DNS Design Flaw Enabling Cloud Infrastructure Takeover

Trend Micro's TrendAI™ research team disclosed a security vulnerability "by design" in the Azure cloud platform. DNS records of deleted Azure resources may persist, allowing attackers to exploit these lingering DNS names to hijack trusted endpoints and compromise dependent systems, highlighting a critical but often overlooked trust inheritance risk in cloud infrastructure.

Palo Alto Networks Other 1970-01-01

Palo Alto's $25B CyberArk Buy Shifts Security Control to Machine Identity

Palo Alto Networks acquires CyberArk for ~$25B to create a unified agent identity and privilege management platform. This shifts security control from network firewalls and EDR to machine identity lifecycle, addressing the AI agent explosion, and ties revenue to token-based consumption.

Google Other 1970-01-01

Google Gemini 3.5 Flash Turns Search into AI-First Answer Engine, Shifting Control from Links to Summaries

Google transforms Search into an AI-first answer engine powered by Gemini 3.5 Flash, with redesigned search bar, AI-generated summary pages, and proactive monitoring. Model improvements include 1M context, 65K output tokens, and multi-agent orchestration via Antigravity, enabling complex task automation.

Research Other 1970-01-01

Z.ai GLM-5.2 Open-Source: 744B MoE, 1M Context, MIT License as Geopolitical Shield

Z.ai releases GLM-5.2: 744B MoE with 40B activated parameters, 1M input and 131K output context, under MIT license. Released one day after Anthropic Fable 5's government takedown, it offers a downloadable, unbanable alternative with Anthropic API compatibility for zero-code migration, giving enterprises a sovereign AI option.

NVIDIA Other 1970-01-01

SGLang 0.5.13 Delivers 25x MoE Inference Speedup via Predictive Routing and Sparse KV Cache

SGLang 0.5.13 introduces two-stage MoE routing prediction and sparse KV cache, achieving a 25x inference speedup on NVIDIA GB300 NVL72. Benchmarks on A100 show 65% throughput gain, 40% latency reduction, and 62% lower routing overhead. This optimization directly attacks the core bottleneck of MoE inference, potentially reshaping AI inference economics.