Youyun Intelligent Computing
Youyun ZhiSuan is a GPU computing power rental platform under UCloud, dedicated to providing a wide range of computing resources for customers in the fields of AI, deep learning, and scientific computing.
Tags:AI programming toolsWhat is Youyun Intelligent Computing?
Compshare, part of UCloud, is a platform for leasing GPU computing power and for accessing APIs related to large-scale models.
It is designed for AI training, model inference, scientific computing, development and testing, as well as production deployment, and offers GPU and model services on demand.
Main functions
- Create GPU instances of different models as needed.
- It supports both containers and virtual machines.
- Hourly, daily, monthly, and preemptive billing are available.
- Charging for GPU computing stops after shutdown.
- The cardless mode can be switched to in order to maintain the operating environment.
- Pre-installed with CUDA, PyTorch, and TensorFlow.
- More than 300 community application images are available.
- Mount the public model library for quick startup.
- Connect via Jupyter, SSH, and VS Code.
- Call the APIs for large language, image, and video models.
- Provides team budget and invoice management.
- Supports API, CLI, and Agent Skills.
Which users are it suitable for
| User | Typical requirements | Select the key points |
|---|---|---|
| AI beginners | Run tutorials and open-source models | Community mirror and hourly fee |
| Algorithm engineer | Training and evaluating models | Video memory, computing power, and data transmission |
| Application developers | Deploy inference services | Stability, concurrency, and networking |
| University teachers | Uniform allocation of student budgets | Teams, bills, and real-name verification |
| Research team | Scientific computing and experimental replication | Mirroring, storage, and task duration |
| Agent users | Use the Coding Plan to invoke the model. | Multiplication factor, concurrency, and cycle quota |
| Corporate clients | Long-term production-level deployment | Region, SLA, and contract terms |
What are the different forms of GPU instances?
| Form | Features | Applicable scenarios |
|---|---|---|
| Container instance | Fast startup, with centralized image environments. | Training, Notebooks, and short-term experiments |
| Virtual machine instance | Features a complete operating system environment. | Custom services and remote desktop |
| Pay-as-you-go instances | Charged based on actual runtime | Unfixed development and reasoning tasks |
| Daily instance | Purchase resources on a daily basis | Intensive experiments carried out over several consecutive days |
| Monthly subscription instances | Long-term retention of computing power | Stable training and continuous support |
| Preemptive instances | The price is lower, but it may be recycled. | Interruptible and resumable tasks |
Common GPU card types
The platform offers GPUs for consumer use as well as for data centers; the availability, memory versions, and stock levels vary depending on the region.
- The RTX 5090 is suitable for new-generation development and testing.
- The RTX 4090 is suitable for training and high-performance inference.
- The RTX 4090 48G is suitable for those with higher memory requirements.
- The RTX 3090 is suitable for experiments with small and medium-sized models.
- The RTX 3080Ti is suitable for tasks where budget is a concern.
- The A100 and A800 are designed for data center workloads.
- H20 is suitable for specific large-model inference scenarios.
- V100S, P40, etc. can be used for compatibility tasks.
How to choose a GPU
| Task | Priority indicators | Suggested method |
|---|---|---|
| Large model fine-tuning | VRAM and training throughput | First, measure the peak memory usage of a single card. |
| Image generation | Video memory and speed per image | Batch testing at fixed resolution |
| Video generation | Video memory and sustained operation capability | Test costs based on target duration |
| Inference service | Concurrency, latency, and stability | Use real traffic for load testing |
| Data processing | CPU, memory, and disk throughput | Don’t focus only on the GPU model. |
| Course experiments | Price and environmental consistency | Unified image and budget limit |
Tutorial on creating GPU instances
- Register and complete the required real-name verification.
- Go to the GPU instance creation page.
- Select the region and availability zone.
- Check the current card type, inventory, and price.
- Determine the number of GPUs, as well as the CPU and memory.
- Choose between container or virtual machine format.
- Choose a basic, community, or custom image.
- Configure the system drive and persistent storage.
- Choose between pay-as-you-go, daily, or monthly billing.
- Create an instance after verifying the estimated costs.
- Wait for the instance to become available.
- Connect via Jupyter, SSH, or VS Code.
- Run small tasks to verify the driver and environment.
- Set up cost reminders and task saving strategies.
How to choose an image?
Base images are suitable for custom environments, while community images are appropriate for rapid deployment of specific frameworks or applications.
| Mirror | Suitable situations | Precautions |
|---|---|---|
| Base image | Install project dependencies manually | Verify the CUDA and framework versions. |
| Community mirror | Quickly launch popular apps | Verify the author, version, and instructions. |
| Custom image | Reuse your full environment | Pay attention to storage costs and sensitive data. |
| Paid images | Use commercial pre-configured solutions | Check for additional fees before creation. |
| Public model library | Avoid re-downloading the model. | Verify the region and instance support scope. |
Methods for connecting to instances
- JupyterLab is suitable for notebook experiments.
- SSH is suitable for terminal management and script tasks.
- VS Code is suitable for remote code development.
- Clients such as FinalShell are suitable for graphical connections.
- Windows virtual machines can use Remote Desktop.
- FileZilla or XFTP can be used for file transfer.
Before making the first connection, verify the instance address, port, username, and the login credentials generated by the platform.
File upload and download
- Choose the transmission method based on the file size.
- Small files can be uploaded through the Jupyter interface.
- Project code can be transmitted via Git or a client.
- Large datasets should be stored in persistent storage first.
- Verify the number and size of files after transmission.
- Independent backups are kept for critical data.
- Confirm that the results have been exported before deleting the instance.
How to understand that there is no charge for shutting down the device?
No charges are applied for GPU computing resources after the instance is shut down, but costs may still apply for resources such as disks, cloud storage, and images.
When it is necessary to preserve the environment but no GPU is to be used, the total cost of the card-free mode can be compared with that of shutting down the system and using persistent storage.
- Save the results promptly after the task is completed.
- Confirm that the process has been completely stopped.
- Check whether the instance is shut down.
- Check the storage resources that are still incurring charges.
- Delete unused images that are no longer valuable after a long period of time.
- Regularly check the bills and the list of resources.
Real-time reference price for GPUs
The following are the lowest prices listed on the official website as of August 31, 2026; location, configuration, inventory levels, and any promotions can affect the final price.
| Resources | Reference price on the official website | Explanation |
|---|---|---|
| RTX 4090 | As low as 2.15 yuan per hour | The inventory levels in Ulanqab, Shanghai, etc. shall prevail. |
| RTX 5090 | As low as 3.32 yuan per hour | Regional and supply dynamics |
| A800 | As low as 6.23 yuan per hour | Suitable for data center workloads |
| Cardless mode | 0.15 yuan per hour | Used to maintain environments that do not have a GPU. |
| Other GPUs | Refer to the console. | Quotations are provided based on card type, region, and configuration. |
Before creation, the final amount shown in the console should be taken as the reference; the system drive, data drive, images, and network may incur additional charges.
Storage price
| Storage type | Free capacity | Expansion or usage price |
|---|---|---|
| Instance cloud disk | 100GB | Expansion cost: 0.01 yuan/GB/day |
| Persistent cloud storage | Based on the account. | 0.004 yuan/GB/day |
| Private image storage | 30GB | Expansion cost: 0.008 yuan/GB/day |
Storage continues to retain data even after the GPU stops working, which can thus become a significant cost for long-term projects.
Large model API
The model API enables developers to invoke models through standard requests, without the need to download, deploy, or maintain the underlying inference environment themselves.
- Supports calling large language models.
- It supports text-to-image and image-to-image generation.
- It supports video generation from text as well as video generation from images.
- It can be connected to Dify and RAGFlow.
- It can be integrated with automation tools such as n8n.
- It can be used with development tools such as Claude Code.
- The specific models, magnifications, and prices will be updated.
Pricing for the Coding Plan package
The Coding Plan provides a quota for model calls on a periodic basis, with different models having varying call multiplicities.
| Package | Price | Approximately 5 hours of invocation time | Called approximately once a week | Called approximately once a month |
|---|---|---|---|---|
| Mini – the compact version | 49 yuan | 300 times | 750 times | 1900 times |
| Lite Starter Edition | 99 yuan | 600 times | 1500 times | 3800 times |
| Basic – Standard version | 199 yuan | 1200 times | 3000 times | 7600 times |
| Pro Enhanced Edition | 499 yuan | 3000 times | 7500 times | 19,000 times |
| Max Premium Edition | 799 yuan | 4800 times | 12,000 times | 31,000 times |
| Ultra Enjoyment Edition | 999 yuan | 6000 times | 15,000 times | 39,000 times |
The numbers in the table are estimates provided by the official website; different tools may send multiple requests for a single task, so the actual number of tasks that can be completed is not equal to the number of requests made.
How to choose a Coding Plan
| Demand | Suggestions | Check the key points |
|---|---|---|
| Short-term experience | Mini or Lite | Target model and scaling factor |
| Personal daily programming | Basic | Weekly and monthly quotas |
| High-frequency development | Pro | Concurrent and single-task call counts |
| Heavy Agent Workflow | Max or Ultra | Toolchain and cost limits |
| Production system | Compare pay-as-you-go APIs | Stability, throttling, and contracts |
Limits on model packages
- Different models have different deduction multiples.
- A single task may result in multiple API calls.
- Different packages have concurrency limits.
- The service will be discontinued once the periodic quota is exhausted.
- Switching to pay-as-you-go may require a new key.
- The list of supported models will be updated by the official team.
- Check the current package’s FAQ before purchasing.
Team and teaching management
The team features allow for inviting members, allocating budgets, recovering unused balances, as well as viewing order and transaction records.
- The root account creates a team and obtains the Team ID.
- Invite registered platform members.
- Wait for team members to accept the team invitation.
- Assign an independent budget to each member.
- Members are asked to switch to the team perspective.
- View member orders and usage details.
- Adjust the amount according to the progress of the course.
- Recover unused budget to the main account.
- Export the necessary bills and operation records.
Cost control for training tasks
- First, validate the code using a small amount of data.
- Monitor GPU utilization and VRAM usage.
- Regularly save checkpoints and logs.
- Interruptible tasks take into account preemptive instances.
- Data preprocessing should not utilize expensive GPUs.
- Shut down immediately after the experiment is completed.
- Clean up unnecessary disks and private images.
- Set up a separate budget for each project.
Data and security considerations
Cloud instances handle code, models, datasets, and keys; businesses and research projects should first determine their data protection requirements.
- Do not store API keys in public repositories.
- Use environment variables or key management methods.
- Restrict SSH ports and login credentials.
- Sensitive data is masked before being uploaded.
- Important results are saved in persistent storage.
- For community mirrors, the source and content are checked first.
- Clean up confidential files before deleting the instance.
- Regulated data must first undergo a compliance assessment.
API, CLI, and open-source status
YouYun Intelligent Computing provides instance management APIs, UCloud SDK calling methods, CompShare CLI, as well as skills designed for agents.
These development tools and examples can be used publicly, but the Youyun AI Cloud platform itself is not an open-source project.
| Project | Open status | Explanation |
|---|---|---|
| Youyun Intelligent Computing Platform | Not open source | Commercial GPU and model cloud services |
| CompShare API | Provide | Manage instances, pricing, and cloud resources |
| UCloud SDK | Accessible | The document provides multilingual ways to make calls. |
| CompShare CLI | Provide | Supports instance, storage, and team management. |
| Agent Skill | Provide | Designed for installation in tools such as Codex. |
| Development example | Public | It does not mean that the source code of the cloud platform is made available. |
API integration steps
- Create an API key in the console.
- Store the public key and private key only in a secure environment.
- Choose the official SDK or request the API directly.
- Set the correct region and availability zone.
- First, call the price and inventory query interface.
- Submit a request to create a minimal instance.
- Record the identifier of the returned instance.
- Polling the instance status and handling failures.
- Stop or delete the resources once the task is completed.
- Fee protection is set for all creation operations.
Product advantages
- There is a wide variety of GPU models and deployment methods available.
- Supports both on-demand and long-term billing options.
- A pre-installed environment reduces configuration costs.
- The public model library reduces duplicate downloads.
- Provides model APIs and Coding Plans.
- The team budget is suitable for teaching and collaboration.
- APIs, CLI, and Skills facilitate automation.
Usage restrictions
- GPU inventory and prices change dynamically.
- Charging may still occur even after the device is turned off.
- The quality of community mirrors needs to be verified by yourself.
- Preemptive instances may be interrupted.
- Model packages have rate and concurrency limits.
- An Agent task may be invoked multiple times.
- Production deployment requires additional stability testing.
Frequently Asked Questions
What services does Youyun Intelligent Computing mainly offer?
It offers on-demand GPU cloud instances, large-model APIs, pre-installed images, a public model library, team budget management, as well as APIs, CLI, and Agent Skills.
How much does YounCloud AI’s GPU cost per hour?
As of August 31, 2026, the official website indicated that the price of the 4090 was as low as 2.15 yuan per unit, the 5090 was as low as 3.32 yuan per unit, and the A800 was as low as 6.23 yuan per hour; the final prices shall be subject to those displayed on the console.
Is there still a charge after Youyun Intelligent Computing is shut down?
After shutdown, charges for GPU computing are usually stopped, but resources such as cloud disks, persistent storage, and private images may still incur fees.
What connection methods does Youyun Intelligent Computing support?
Containers can be accessed via Jupyter, SSH, FinalShell, and VS Code; Windows virtual machines also support remote desktop access. The specific methods of access depend on the type of instance.
Is the number of times the Coding Plan is called equal to the number of tasks?
It’s not the same; an Agent or a programming task may send multiple requests to the models, and different models have different scaling factors, so the actual number of tasks varies.
Does Youyun Intelligent Computing provide APIs and CLI?
Developers can use the CompShare API, UCloud SDK, CompShare CLI, and Agent Skill to manage instances, storage, and team resources.
Is Youyun Zhi Suo an open-source platform?
No, the cloud platform itself is a commercial service; the availability of the official CLI, Skills, and development examples does not mean that the backend source code of the platform is made public.
Guigong Network Security Registration No. 45132202000164