Key article takeaways

  • GreenNode launched the HAN-1B Availability Zone in Hanoi to expand its AI cloud capacity and regional coverage.
  • The new zone adds more high-performance CPU/GPU resources for scalable, low-latency AI workloads.
  • This expansion advances GreenNode’s strategy to become a leading AI cloud provider in Vietnam and Southeast Asia.

GreenNode has officially launched a new Availability Zone, HAN-1B, in Hanoi, expanding its AI Cloud infrastructure capacity and enabling enterprises to deploy systems with high availability, stable performance, and flexible scalability.

The addition of Availability Zone HAN-1B marks an important milestone in GreenNode’s journey to build a hyperscale-class AI Cloud platform in Vietnam, supporting modern workloads such as AI/ML, Big Data, Kubernetes, and large-scale backend systems for enterprises across industries including finance and banking, retail, and media.

GreenNode expands the HAN Region with new Availability Zone HAN-1B

Previously, GreenNode’s Hanoi region operated with a single availability zone, HAN-1A. With the launch of HAN-1B, the HAN Region now consists of two independent Availability Zones.

This architecture enables customers to deploy systems using a multi-AZ cloud architecture, improving application resilience and fault tolerance. It also helps enterprises meet data storage and protection compliance requirements under Vietnamese regulations such as the Data Law 2024 and the Personal Data Protection Law 2025.

Following the addition of Availability Zone HAN-1B, GreenNode’s infrastructure now includes:

  • HAN Region (Hanoi): 2 Availability Zones
  • HCM Region (Ho Chi Minh City): 3 Availability Zones
  • BKK Region (Bangkok): 1 Availability Zone

For industries with strict security and compliance requirements, such as finance and banking, where transaction systems and customer data must operate 24/7, a multi-region and multi-Availability Zone cloud architecture plays a critical role in ensuring both system availability and regulatory compliance.

Post 1_New AZ_email (1).png

Key infrastructure highlights

  • Network backbone connectivity up to 50 Gbps
  • Low latency between Availability Zones (<5 ms)
  • Multi-region infrastructure across Hanoi, Ho Chi Minh City, and Bangkok
  • Up to 99.99% SLA when deploying Multi-AZ architecture.
greennode-blog-new-cloud-instance-pic-1-en.png
The infrastructure meets international standards, allowing customers to easily deploy a multi-AZ architecture similar to a hyperscaler to increase system availability.

In addition, GreenNode provides connectivity and security solutions that help enterprises build a more secure infrastructure architecture:

  • VPC can be deployed across two Availability Zones within the same region.
  • Private Endpoint ensures that traffic between the VPC and cloud services remains within the private network, minimizing the need to traverse the public internet and significantly reducing the attack surface.
  • VPC Peering: Enables private connectivity between VPCs across different zones, allowing traffic routing without exposing services to the internet.
  • VPN, Point-to-Point (P2P), and Interconnect provide dedicated connections between on-premise systems and the cloud, improving security and stability compared to public internet connections.
  • Multi-AZ architecture ensures high availability and continuous system operation, including monitoring and security systems.

This expansion further strengthens GreenNode’s AI Cloud platform across Hanoi, Ho Chi Minh City, and Bangkok, enabling enterprises to deploy flexible infrastructure for AI workloads and mission-critical systems.

Cloud services launched with Availability Zone HAN-1B

Along with the launch of Availability Zone HAN-1B, GreenNode is also deploying a comprehensive suite of cloud core services to enable businesses to quickly operate their systems.

Services to be deployed with HAN-1B in the near future include:

  • Compute & Storage: Virtual Machine (VM), Block Storage/Volume, Backup.
  • Networking: Load Balancer, NAT Gateway, VPN, Private Endpoint, Bandwidth.
  • Cloud Platform Services: Marketplace, VKS (Managed Kubernetes Service).

In addition, the self-service network on the portal allows users to easily configure and manage the infrastructure according to their needs.

Strengthening AI GPU Infrastructure in Hanoi

Alongside the Availability Zone expansion, GreenNode is also enhancing its AI GPU infrastructure capacity in the Hanoi region.

The platform currently supports high-performance GPU models such as NVIDIA H100 Tensor Core GPU and NVIDIA A40 GPU. Combined with a multi-AZ architecture, AI workloads can achieve high performance, strong stability, and flexible scalability.

Below are the recommended workloads for each GPU line offered by GreenNode, particularly for industries with strict security requirements and intensive AI usage suck as BFSI:

GPU Model

BFSI Use cases

Workload Category

Production Notes

NVIDIA H100 GPU

Advanced ML:

  • Real-time market making with microsecond-latency neural networks.
  • Distributed training of large fraud detection ensembles.
  • Federated learning across multiple bank branches.

Advanced Computer Vision:

  • Real-time video KYC with liveness detection.
  • Multi-modal fusion (text + image + tabular) for underwriting.
  • Synthetic data generation (GANs, SDXL) for rare fraud patterns.

Large LLM & GenAI:

  • Fine-tuning Llama 3 70B on proprietary financial data.
  • Training custom 30B-70B domain-specific models.
  • Multi-agent orchestration: compliance, risk, and audit agents working together.
  • RAG at scale with real-time knowledge graph updates.
  • Code generation for automated trading strategies. 
Production GenAI & Training
  • Gold standard for LLM inference and fine-tuning.
  • NVLink, InfiniBand for multi-GPU, multi-node scaling.
  • Essential for training custom foundation models. 
NVIDIA A40 GPU

Traditional AI:

  • Legacy model inference (TensorFlow 1.x, older PyTorch models).
  • Large-scale batch scoring for credit applications.
  • Time-series forecasting (LSTM/GRU) for market prediction.

Computer Vision:

  • Batch processing of mortgage document scans.
  • Large-scale image classification for insurance claims.
  • Video surveillance analytics (non-real-time).

LLM Workloads:

  • BERT/RoBERTa fine-tuning for regulatory text classification.
  • Smaller LLM inference (up to 13B with int8 quantization).
  • Embedding generation for semantic search in knowledge bases. 
Legacy Workload & Batch Processing
  • Still widely deployed.
  • Good for organizations with established Ampere workflows.
  • Lower power consumption. 
NVIDIA L40S GPU

Traditional AI:

  • Batch fraud detection across millions of transactions.
  • Portfolio optimization with deep reinforcement learning.
  • Large-scale credit risk modeling with neural networks.

Computer Vision:

  • Video analytics for branch security (multi-stream processing).
  • Automated underwriting with multi-page document analysis.
  • Insurance telematics analysis (dashcam footage).

Production LLM:

  • Llama 3 70B inference (with quantization) for customer support.
  • Fine-tuning 13B-30B models for compliance document generation.
  • Multi-agent systems: fraud investigation with specialized AI agents. 
Production Inference & Multi-User Environments
  • Enterprise-grade reliability.
  • Excellent for shared inference servers.
  • Good balance of VRAM and cost.
  • Passive cooling for data center.
NVIDIA GeForce RTX 5090

Traditional AI:

  • Real-time risk assessment dashboards with ensemble models.
  • Market microstructure analysis with deep learning.
  • Customer churn prediction with neural networks.

Computer Vision:

  • Multi-modal document understanding (receipts, invoices, forms).
  • Biometric authentication (face + voice fusion).
  • Property damage assessment for insurance.

Medium LLM:

  • Fine-tuning Llama 3 8B/13B for financial report summarization.
  • RAG systems with embedding models for policy Q&A.
  • Agentic workflow: automated loan pre-screening with LLM reasoning. 
Advanced Development & Edge Deployment
  • Excellent for edge AI deployments (branch offices).
  • Supports larger model fine-tuning. Better thermal design than NVIDIA GeForce RTX 4090.
NVIDIA GeForce RTX 4090

Traditional AI:  

  • Credit scoring models (XGBoost, LightGBM inference).
  • Fraud detection with CNN-based transaction pattern analysis.
  • Document OCR for KYC/AML (Tesseract, EasyOCR).

Computer Vision:

  • Check image processing and signature verification.
  • ATM surveillance anomaly detection.
  • Insurance claim photo assessment.

Small LLM/NLP:

  • Fine-tuned BERT/DistilBERT for sentiment analysis on customer feedback.
  • Small chatbot models (7B params) for customer service
  • Named Entity Recognition for contract analysis. 
Development & Small Production
  • Great for PoC/Development. Single-user workstations.
  • Not ideal for 24/7 production due to consumer-grade cooling. 
NVIDIA H200 GPU (coming soon)*

Cutting-Edge ML:

  • Ensemble of ensembles: meta-learning across all risk models.
  • Graph neural networks for complex financial network analysis.
  • Multi-objective optimization for portfolio management.

Next-Gen Computer Vision:

  • End-to-end document intelligence (OCR + Understanding + Extraction).
  • Real-time video analysis for physical security + transaction correlation.
  • Generative models for synthetic financial scenario creation.

Frontier LLM & Agentic AI:

  • H200 Running Kimi K2.5 with 141GB HBM3e VRAM and 4.8 TB/s bandwidth to leverage Ultra-Long Context (2M+ tokens) for End-to-End Financial Auditing and Multi-year Compliance Review.
  • Multi-agent orchestration: 10+ specialized financial agents.
  • Training custom 100B+ parameter models on proprietary data.
  • Continuous fine-tuning pipelines with real-time market data.
  • Agentic workflows: end-to-end loan processing (application → underwriting → approval).
  • Multi-modal foundation models (text + time-series + images). 
Frontier AI & Research
  • Maximum VRAM for largest models.
  • Ideal for organizations building proprietary foundation models.
  • Critical for long-context applications. 

AI Cloud infrastructure for enterprise workloads

GreenNode is designed to support workloads that demand high compute performance, strong networking, and system stability.

Common workloads include:

  • AI / Machine Learning pipelines
  • Large-scale inference systems
  • Big Data analytics (Spark, ClickHouse, Presto)
  • Backend services with high traffic volumes
  • Recommendation systems and real-time analytics

With a high-speed network backbone, powerful GPUs, and a multi-AZ architecture, GreenNode enables enterprises to build modern AI infrastructure within Vietnam. At the same time, the platform implements comprehensive data protection practices across the entire Cloud Platform, ensuring peace of mind for customers and supporting enterprises in addressing questions related to new data regulations.

The launch of Availability Zone HAN-1B in Hanoi marks an important milestone for GreenNode in expanding its high-performance AI Cloud platform for enterprises, enabling organizations to deploy AI systems, data platforms, and mission-critical applications with high availability while keeping infrastructure located in Vietnam.

Contact the GreenNode team today to receive consultation on the optimal infrastructure architecture for your enterprise applications.