< ciso
brief />
Tag Banner

All news with #product launch tag

547 articles · page 17 of 28

Amazon EC2 M4 Max Mac Instances Now Generally Available

🚀 Amazon Web Services announced general availability of Amazon EC2 M4 Max Mac instances, powered by the latest Mac Studio hardware. These next‑generation Mac instances deliver up to 25% better application build performance versus Amazon EC2 M1 Ultra Mac and target demanding build and test workloads for Apple platforms. They run on Apple M4 Max silicon (16‑core CPU, 40‑core GPU, 16‑core Neural Engine, 128 GB unified memory), use the AWS Nitro System, and offer up to 10 Gbps network and 8 Gbps EBS bandwidth; availability begins in US East (N. Virginia) and US West (Oregon).
read more →

Amazon EC2 C8i Instances Now Available in London Region

🚀 Amazon EC2 C8i instances are now available in the Europe (London) region, powered by custom Intel Xeon 6 processors exclusive to AWS. They deliver up to 15% better price-performance and 2.5x more memory bandwidth versus previous-generation Intel-based instances, and up to 20% higher performance compared with C7i. AWS cites workload-specific gains — up to 60% faster NGINX, 40% faster AI recommendation models, and 35% faster Memcached — and offers 13 sizes including two bare metal options and a new 96xlarge. Customers can purchase via Savings Plans, On-Demand, or Spot instances.
read more →

Amazon EC2 C8i and C8i-flex Now in Sydney and Frankfurt

🚀 Amazon Web Services has added EC2 C8i and C8i‑flex instances to the Asia Pacific (Sydney) and Europe (Frankfurt) regions. These instances are built on custom Intel Xeon 6 processors available only on AWS and deliver up to 15% better price-performance and 2.5x greater memory bandwidth versus prior Intel-based generations. AWS cites up to 20% higher overall performance versus C7i families and workload-specific gains—up to 60% for NGINX, 40% for deep-learning recommendation models, and 35% for Memcached. C8i‑flex targets common compute sizes (large to 16xlarge) while C8i targets memory-intensive use cases with 13 sizes including 96xlarge and bare-metal; customers can buy On-Demand, Spot, or via Savings Plans.
read more →

Amazon Neptune Analytics Generally Available in New Regions

🚀 Amazon Neptune Analytics is now generally available in additional AWS regions, including US West (N. California), Asia Pacific (Seoul, Osaka, Hong Kong), Europe (Paris, Stockholm), and South America (São Paulo). The serverless Amazon Neptune graph database scales automatically and supports advanced graph analytics and fully managed GraphRAG capabilities. Neptune models data as a graph to capture context that improves accuracy and explainability for AI applications, and integrates with Amazon Bedrock, Strands AI Agents SDK, and common agentic memory tools. Create and manage Neptune Analytics graphs via the AWS Management Console or AWS CLI; refer to the Neptune pricing page and AWS Region Table for pricing and availability.
read more →

AlloyDB Introduces Managed Connection Pooling for PostgreSQL

🔌 AlloyDB now offers managed connection pooling for PostgreSQL as a generally available capability, reducing connection overhead and improving scalability. The service-managed pooler listens on port 6432 and reuses backend connections to cut latency and resource churn, while Google Cloud handles setup, configuration, and maintenance. Choose Transaction or Session pooling based on compatibility needs and configure sizes via Console, gcloud, or API. Customers report support for 3x more clients and up to 5x higher transactional throughput.
read more →

Scaling MoE Inference with NVIDIA Dynamo on A4X Rack-Scale

🚀 This post describes two validated deployment recipes for serving large Mixture-of-Experts (MoE) models on Google Cloud's A4X machines using NVIDIA Dynamo. The recipes provide throughput- and latency-optimized configurations that exploit the 72‑GPU GB200 NVL72 rack fabric, WideEP/DeepEP parallelism, global KV cache, and GKE-aware rack-level scheduling. Performance validation reports >6K tokens/sec/GPU for the throughput recipe and a 10ms median inter-token latency for the latency-optimized recipe.
read more →

Office, Visio and Project 2024 on Amazon WorkSpaces

🖥️ Amazon WorkSpaces Personal and Core now include Microsoft Office, Visio, and Project 2024 applications in the managed applications catalog. The release covers Microsoft Office LTSC Professional Plus 2024, Office LTSC Standard 2024, Visio LTSC Professional and Standard 2024, and Project Professional and Standard 2024. Administrators can add these applications to eligible new or existing WorkSpaces using the existing Manage application workflow to standardize a modern, secure productivity desktop. Applications are available in all Regions that support WorkSpaces Personal and Core and will incur per-application charges.
read more →

AWS Graviton4 EBS-Optimized C8gb/M8gb/R8gb 48xlarge

🚀 AWS has made EBS-optimized EC2 C8gb, M8gb, and R8gb instances in 48xlarge sizes generally available, with C8gb and R8gb also offered as metal-48xl. Powered by Graviton4, the new sizes deliver up to 30% better compute performance than Graviton3 and offer up to 300 Gbps of EBS bandwidth and 1,440,000 IOPS. They support Elastic Fabric Adapter (EFA) and up to 400 Gbps networking to improve latency and cluster performance for tightly coupled and high-throughput workloads. New sizes are available in US East (N. Virginia) and US West (Oregon).
read more →

Amazon EC2 G7e Instances Now GA with NVIDIA Blackwell

🚀 Amazon EC2 G7e instances are now generally available, powered by NVIDIA RTX PRO 6000 Blackwell Server Edition GPUs. G7e delivers up to 2.3x inference performance versus G6e and supports configurations with up to 8 GPUs (96 GB each), 5th Gen Intel Xeon processors, 192 vCPUs, and up to 1600 Gbps of Elastic Fabric Adapter networking. Designed for LLMs, multimodal and spatial computing workloads, G7e includes NVIDIA GPUDirect P2P and RDMA support in EC2 UltraClusters and is available in US East (N. Virginia) and US East (Ohio) as On‑Demand, Spot, or via Savings Plans.
read more →

Getting Started with Gemini 3 Flash on Google Cloud

🚀 This post introduces Gemini 3 Flash, Google’s low-latency, cost-efficient model in the Gemini 3 family, optimized for advanced reasoning, multimodal understanding, and agentic workflows. It guides developers through obtaining an API key from Google AI Studio and configuring it for local use or environment-based invocation. The article demonstrates interactive prompt testing in the Playground, explains toggles like Structured outputs and Thinking level, and shows how to export language-specific sample code via the "Get code" feature to run with the Google GenAI SDK.
read more →

OpenAI Offers One-Month Free ChatGPT Plus Subscription

🔔 OpenAI is offering a free one-month trial of ChatGPT Plus, normally $20/month, through a limited-time promotion available to many accounts. The offer can be activated now and canceled anytime before it auto-renews, so users who want to avoid charges must cancel before the end of the month. Plus provides higher message and file limits, expanded memory, and longer context windows than the free or Go tiers. OpenAI also plans to introduce ads into the Free and Go tiers in the coming weeks.
read more →

OpenAI Hostname Suggests New ChatGPT Feature 'Sonata'

🎵 OpenAI has started using new hostnames—sonata.openai.com and sonata.api.openai.com—spotted on 15–16 January 2026, suggesting work on a service codenamed Sonata. A new subdomain typically signals a web-facing product, internal tool, or API, but the codename alone doesn't confirm functionality. OpenAI recently improved ChatGPT's reference chat history retrieval and expanded dictation, which could align with audio or transcription enhancements.
read more →

OpenAI launches ChatGPT Go worldwide at $8 with ads

🔔 OpenAI has rolled out the $8 ChatGPT Go subscription globally, offering users 10× more messages, increased file uploads, expanded image creation, longer memory, and a larger context window than the free tier. Go provides access to the latest GPT-5.2 Instant but does not include the higher-tier "reasoning" models reserved for paid plans. The Go tier displays ads; upgrading to GPT Plus ($20) or GPT Pro ($200) removes them and restores advanced model access.
read more →

Astro Joins Cloudflare to Accelerate Web Development

🚀 Cloudflare has acquired The Astro Technology Company and will integrate the Astro web framework into its platform while keeping the project open source under the MIT license. All full-time Astro employees have joined Cloudflare and the company pledges continued support for the Astro Ecosystem Fund alongside partners. Astro 6 is in public beta, featuring a redesigned development server built on the Vite Environments API, stable Live Content Collections, improved Content Security Policy support, and simpler APIs.
read more →

Microsoft's Copilot Studio VS Code Extension Public

🚀 Microsoft released the Copilot Studio extension for Visual Studio Code, enabling developers to build and manage Copilot Studio agents directly within the editor. The extension lets teams pull full agent definitions locally, edit components with IDE features like syntax highlighting and IntelliSense-style completion, and preview or compare changes against the cloud. It supports Git versioning, CI/CD integration, and works with AI coding assistants to speed development; the extension is free on the VS Code Marketplace and has been downloaded over 13,000 times.
read more →

AWS Databases Now Available in Vercel's v0 Environment

🚀 Amazon Aurora PostgreSQL, Amazon Aurora DSQL, and Amazon DynamoDB serverless databases are now accessible directly from v0 by Vercel, letting developers build full-stack applications and connect to AWS databases using natural language prompts. v0 provides an end-to-end setup experience to create or link AWS accounts, with new accounts receiving access to all three databases and $100 USD in credits. Serverless options scale to zero and are available in seven AWS Regions, reducing operational overhead for prototypes and production AI-driven applications.
read more →

BigQuery: Managed SQL-native Inference for Open Models

🚀 BigQuery now supports managed third‑party generative AI inference (Preview) for open models from Hugging Face and Vertex AI Model Garden, enabling SQL-native deployment and inference. With a single CREATE MODEL statement you can provision and configure compute, control lifecycle with endpoint_idle_ttl and ALTER MODEL, and run inference via AI.GENERATE_TEXT or AI.GENERATE_EMBEDDING. BigQuery automates resource cleanup and integrates cost controls to reduce operational overhead.
read more →

Design an AI and Agent Strategy with Microsoft Marketplace

🧭 Microsoft Marketplace positions itself as the central catalog for organizations choosing how to adopt AI—whether to build, buy, or blend solutions. It hosts more than 11,000 prepackaged models and over 4,000 AI apps and agents, accessible via the storefront, Azure portal, and Microsoft Foundry. The platform supports both pro-code and low-code development workflows, including Copilot Studio, and emphasizes integration, governance, and faster time-to-value for enterprise deployments.
read more →

AWS Launches EC2 X8i: Next-Gen Memory-Optimized Instances

🚀 AWS has announced general availability of Amazon EC2 X8i instances, a new family of memory-optimized instances built on custom Intel Xeon 6 processors exclusive to AWS. X8i offers up to 43% higher overall performance, 1.5x more memory capacity (up to 6 TB) and 3.4x greater memory bandwidth compared to prior X2i instances. Designed for SAP HANA, large databases, data analytics and EDA, X8i is SAP-certified and available in 14 sizes including two bare-metal options. Instances are initially available in US East (N. Virginia), US East (Ohio), US West (Oregon) and Europe (Frankfurt) and can be purchased via Savings Plans, On-Demand or Spot.
read more →

OpenAI's Hidden ChatGPT Translate Rivals Google Translate

🌐 OpenAI has quietly launched ChatGPT Translate, a web-based translation tool accessible at chatgpt.com/translate and available to all users without a paid account. It supports typed text, photo uploads, voice input, and file attachments, automatically detecting language or allowing manual source/target selection. The tool emphasizes preserving meaning over literal translations and lets users request tones like “business formal” or “explain like a child,” with the added benefit of continuing the conversation to refine results. ChatGPT’s Android and iOS apps do not yet expose the translate toggle.
read more →