Backend.AI dev environments Enables seamless customization and hacking of Backend.AI Empowers everyone to own and modify their AI infrastructure bndev https://bnd.ai/bndev
NemoTron-4-340B Starts from Llama3.1, Gemma2 Built on FastTrack Runs on Backend.AI Cloud The waitlist is now open! finetun.ing https://www.finetun.ing
scale and performance Automatically calculates effective performance, required hardware, and estimated costs Ideal for validating optimal architecture before actual deployment Design your own cluster at our demo booth! Backend.AI Cluster designer
Inference runtime Hugging Face model via model URL PALI: Performant AI Launcher for Inference Model Store Open Models Lablup GPU Virtualizer Backend.AI Model Player Partner Models
scalable by connecting multiple PALI-equipped appliances Optimized architecture for AI workloads – Delivers high performance and low latency PALI Performant AI Launcher for Inference Intel Gaudi Gaudi 2/3 integration GraceHopper GH200 / GB200 A6000 / L40 x86-64 based node Model Store Per architecture / chip PALI2: Scalable AI H/W Infrastructure
on NVIDIA GH200 reference platform (Korea) – Pre-orders in October, sales from Q4 Instant.AI by Kyocera Mirai Envision (Japan) – Launching on October 1st, 2024 PALI^2 appliances for US and European markets – Expected as early as Q4 this year https://www.kcme.jp/product/instant-ai-server/
Helmsman Simplifies large-scale language model deployment and operation – Provides ready-to-use inference and fine- tuning settings Talkativot enables easy creation of customized chatbots PALANG: PALI for LANGuage models
Helmsman Simplifies large-scale language model deployment and operation – Provides ready-to-use inference and fine- tuning settings Talkativot enables easy creation of customized chatbots PALANG: PALI for LANGuage models PALI Performant AI Launcher for Inference GraceHopper GH200 / GB200 A6000 / L40 x86-64 based node Model Store Per architecture / chip FastTrack MLOps 2 Helmsman Conversional Backend.AI management UX Talkativot Chatbot UI for Language models LLMs Model weights H100 / B200 x86-64 based node
Model Recpies for AI Inference BNDEV DevStack manager FastTrack MLOps 2 CLI Installer Interactive terminal UI Site Designer Helmsman Conversional Backend.AI management UX PALI Performant AI Launcher for Inference PALI2 PALI Appliance PALANG Language model-oriented AI Inference platform GARNET Gemma2-based SLM Next-gen Sokovan Also with Kubernetes finetun.ing Model tuning by discussion, without data WebUI 3 Neo Rewritten UI/UX