This message was deleted.
# ask-for-help
s
This message was deleted.
b
@rekil You should able to use
terraform destroy
command to do that
r
thank you very much will try that , after that i should generate the terraform with -
bentoctl generate -f deployment_config.yaml
and continue correct?
b
yep
r
thanks i'll try and report back
sorry but what do i put in here
b
you could use
terraform destroy -var-file bentoctl.tfvars
πŸ‘ 1
r
ok successfully deleted, let me rerun the other commands
hmm back to the same error :
Copy code
β•·
β”‚ Error: creating IAM instance profile stable-diffusion-bentoml-instance-profile: EntityAlreadyExists: Instance Profile stable-diffusion-bentoml-instance-profile already exists.
β”‚ 	status code: 409, request id: be742c25-c728-42fc-8e8f-bcc78863894f
β”‚ 
β”‚   with aws_iam_instance_profile.ip,
β”‚   on <http://main.tf|main.tf> line 90, in resource "aws_iam_instance_profile" "ip":
β”‚   90: resource "aws_iam_instance_profile" "ip" {
β”‚ 
β•΅
β•·
β”‚ Error: creating Security Group (stable-diffusion-bentoml-bentoml-sg): InvalidGroup.Duplicate: The security group 'stable-diffusion-bentoml-bentoml-sg' already exists for VPC 'vpc-a96110d1'
β”‚ 	status code: 400, request id: 724408a5-52c5-44b5-a914-26dc31064575
β”‚ 
β”‚   with aws_security_group.allow_bentoml,
β”‚   on <http://main.tf|main.tf> line 96, in resource "aws_security_group" "allow_bentoml":
β”‚   96: resource "aws_security_group" "allow_bentoml" {
β”‚ 
β•΅
b
hmmm
those should deleted by terraform
can you manually delete them and try again?
also have to hop on a call for 20 mins to debug
r
ok thank you, i can try and get back
so managed to delete those but i'm back to the vCPU limit error, i feel i'm missing something very basic:
Copy code
β•·
β”‚ Error: creating EC2 Instance: VcpuLimitExceeded: You have requested more vCPU capacity than your current vCPU limit of 32 allows for the instance bucket that the specified instance type belongs to. Please visit <http://aws.amazon.com/contact-us/ec2-request> to request an adjustment to this limit.
β”‚ 	status code: 400, request id: 50a70b98-b99a-4103-84c1-8e40a8e13a8e
β”‚ 
β”‚   with aws_instance.app_server,
β”‚   on <http://main.tf|main.tf> line 141, in resource "aws_instance" "app_server":
β”‚  141: resource "aws_instance" "app_server" {
s
What size of
g4dn
node are you requesting,
xlarge
or
2xlarge
?
Do you have any existing allocated nodes of
g4dn
in use in the region of your deployment?
@rekil any update here?
r
@Sean i tried on xlarge, 2xlarge andd 4xlarge.
but it's giving the vCPU limit error.
s
Do you have any existing running instances of GPU nodes in the region?
r
i ssh into that machine and follow the instructions on the github page. I feel i'm missing something obvious
s
This error is thrown from AWS when running terraform against the EC2 provider. EC2 instance cannot be allocated due to exceeding vCPU quota. A single
g4dn.xlarge
consumes 4 quota units. Quotas are also per region. I’d verify if there are other GPU enabled instances running in the same region.
Have you tried requesting a
g4dn.xlarge
instance directly from EC2? Do we get the same error?
r
So we increased the vCPU limit to 32 and tried with a g4dn.2xlarge but still got the same error on limits. Is there someway to set up a call to run through this error, just to make sure i'm not missing something very obvious