Frequently Asked Questions

Product Information & VM Optimization

What is Sedai and how does it optimize virtual machines (VMs) in the cloud?

Sedai is an autonomous cloud platform that optimizes cloud operations—including virtual machines (VMs)—using advanced machine learning and AI. For VM-based applications, Sedai identifies optimization opportunities, executes autonomous remediations, and manages performance and cost by right-sizing resources, scheduling selective shutdowns, and validating every action for safety. Sedai supports VM optimization across AWS, Azure, VMware, and on-premise environments, with GCP support planned by the end of Q4. Note: GCP support is not yet available; check with Sedai for the latest rollout status.

Which cloud platforms and environments does Sedai support for VM optimization?

Sedai currently supports VM-based applications on AWS, Azure, VMware, and on-premise environments (via custom integrations). Support for Google Cloud Platform (GCP) is planned for the end of Q4. Note: GCP support is not yet available; confirm with Sedai for current status.

How does Sedai ensure safe, autonomous optimization for VMs?

Sedai is designed with safety as a core principle. Every autonomous optimization action is validated through continuous health checks, automatic rollbacks, and incremental changes. Before executing any remediation or optimization, Sedai assesses risk, checks timing, and validates application health post-action. This approach minimizes the risk of outages or SLO breaches. Note: Detailed limitations not publicly documented; ask sales for specifics on edge cases or highly customized VM environments.

What are the main challenges of managing VMs that Sedai addresses?

Sedai addresses common VM management challenges such as right-sizing and capacity planning, provisioning and configuration, security and compliance, patch management, cost optimization, backup and disaster recovery, automation, orchestration, monitoring, and diagnostics. By automating these tasks and providing actionable insights, Sedai reduces operational toil and risk for SREs and DevOps teams. Note: For highly specialized compliance or legacy requirements, manual oversight may still be necessary.

Features & Capabilities

What autonomous actions does Sedai perform for VM availability and optimization?

Sedai autonomously executes remediations based on detected anomalies in latency, errors, and custom metrics. It recommends and applies optimal auto-scaling configurations, vertical resizing, and selective shutdowns for development and QA instances based on predicted traffic and seasonality. All actions are validated for safety and efficacy. Note: Some actions may require manual approval in highly regulated or complex environments.

How does Sedai handle non-standard VM setups or multiple applications on a single VM?

Sedai identifies multiple applications deployed on a single VM by analyzing port usage and associated metrics. It schedules operations during low-traffic periods or planned downtimes to minimize disruption. Note: Handling highly customized or legacy VM setups may require additional manual configuration or oversight.

What integrations does Sedai offer for VM monitoring and automation?

Sedai integrates with 12 APMs (including Prometheus, Datadog, AWS CloudWatch, Azure Monitor, and Google Cloud Monitoring), Kubernetes autoscalers (HPA/VPA, Karpenter), IaC and CI/CD tools (GitHub, GitLab, Bitbucket, Terraform), ITSM tools (ServiceNow, PagerDuty, Jira), notification platforms, and runbook automation systems. This enables comprehensive monitoring and automation for VM environments. Note: Some integrations may require additional setup or configuration depending on your environment.

Implementation & Ease of Use

How long does it take to implement Sedai for VM optimization, and how easy is it to get started?

Initial setup for Sedai can be completed in as little as 15 minutes using agentless or agent-based deployment. For more advanced AI Agent Optimization, implementation typically takes two to three weeks. Customers report quick onboarding, seamless integration with existing tools, and minimal management burden due to Sedai's autonomous operation. Note: Implementation time may vary for highly customized or legacy VM environments.

What technical documentation and support resources are available for Sedai VM optimization?

Sedai provides comprehensive onboarding guides, technical documentation for VM optimization, and step-by-step instructions for integrating with various environments. Resources are available at docs.sedai.io/get-started. Personalized onboarding, a community Slack channel, and real-time support are also available. Note: Some advanced scenarios may require direct engagement with Sedai's support team.

Pricing & Plans

How is Sedai priced for VM optimization?

Sedai uses a resource-based pricing model, where costs are determined by the resources optimized and the value delivered. All costs are transparently outlined on Sedai's pricing page, with no hidden fees. Discounts from connected cloud billing accounts (e.g., Reserved Instances, Savings Plans) are factored into cost and savings calculations. Note: For specific pricing details for your VM environment, contact Sedai sales or request a demo.

Security & Compliance

What security and compliance certifications does Sedai have?

Sedai is SOC 2 certified, demonstrating adherence to stringent security and data protection standards. This certification ensures compliance with industry requirements for cloud operations. For more details, visit Sedai's Security page. Note: For additional certifications or compliance requirements, contact Sedai directly.

Use Cases & Customer Success

What types of organizations and roles benefit most from Sedai's VM optimization?

Sedai's VM optimization is valuable for IT/cloud operations managers, SREs, DevOps teams, FinOps leads, technology leaders (CTO, CIO, VP Engineering), and platform engineers in organizations managing hybrid or multi-cloud environments. It is especially useful for teams seeking to reduce operational toil, optimize costs, and improve application performance. Note: Organizations with highly specialized legacy systems may require additional integration work.

What business impact and performance improvements can Sedai deliver for VM environments?

Sedai delivers up to 50% reduction in cloud costs by rightsizing workloads and eliminating cloud waste, up to 75% reduction in application latency, and up to 6X productivity gains for engineering teams by automating repetitive tasks. These outcomes are based on customer case studies and platform metrics. Note: Actual results may vary depending on environment complexity and baseline efficiency.

Can you share specific customer success stories related to VM optimization?

Yes. For example, Palo Alto Networks saved $3.5 million through Sedai's autonomous optimization, and KnowBe4 achieved up to 50% cost savings and a 99.5% reduction in response time for AWS Lambda workloads. For more details, see the KnowBe4 case study and Palo Alto Networks video case study. Note: Results are specific to each environment; VM-specific outcomes may vary.

Introducing Sed: Your cloud & AI assistant

Meet Sed
Sedai Logo

Virtual Machines in the Autonomous Era: New Optimization Strategies for SREs and DevOps

Aby Jacob Headshot

Aby Jacob

VP of Engineering

October 14, 2024

In cloud computing, virtual machines (VMs) play a crucial role in modern infrastructure. Despite the rise of alternatives like Kubernetes and serverless platforms, many organizations still rely on VMs for their flexibility, fine-grained control, and compatibility with hybrid environments. VMs are also helpful in ensuring compliance with standards and providing better security and isolation. 

However, managing a large fleet of VMs can present challenges, especially for SREs and DevOps teams. This article explores how platform teams can leverage autonomous optimization strategies to streamline VM management and improve overall operational efficiency.

Significance of VMs in Modern Infrastructure

Virtual machines (VMs) have been the foundation of cloud computing for decades. Even with the emergence of modern platforms like Kubernetes and serverless architectures, VMs remain widely used. Here’s why they continue to hold their ground in today’s cloud infrastructure:

  • Flexibility and Control: VMs offer fine-grained control, which is particularly valued by large enterprises. This level of control allows companies to tailor their infrastructure to specific requirements.
  • Compliance and Security: Many organizations prefer VMs because they provide strong security measures and help maintain compliance with industry standards. Their isolated nature makes them a solid choice for secure operations.
  • Hybrid Cloud Compatibility: VMs seamlessly support hybrid cloud environments, making it easier for companies to move workloads between on-premise systems and the cloud. This mobility offers a practical solution for businesses transitioning to cloud-based systems.
  • Cost Efficiency: VMs offer an advantage in terms of cost savings for predictable workloads. Businesses can lower their infrastructure expenses by right-sizing applications and optimizing resource usage.
  • Versatility for Various Workloads: VMs are ideal for handling a wide range of workloads, from stateless web applications to stateful databases, high-performance computing, and batch processing. This versatility makes them an essential tool for diverse computing needs.

A key comparison is drawn between modern IaaS (Infrastructure as a Service) and VMs:

  • IaaS (Bentley analogy): Expensive but comes with standardized practices that streamline operations.
  • VMs (Land Cruiser analogy): Rugged, customizable, and versatile, capable of handling complex, tailored workloads much like a workhorse vehicle.

Having a modern IaaS might not solve all your infrastructure problems. For example, a whitepaper published by Amazon Prime highlighted the cost benefits of VMs. When they switched from a serverless and step-function-based application to one using EC2 and ECS tasks, they achieved the same quality of results but at a fraction of the cost. 

Even though VMs may seem like legacy technology, they remain highly relevant. As the saying goes, "An old broom knows the dirty corners best." VMs, while traditional, offer reliability and proven performance in many enterprise scenarios.

With increased control comes the responsibility of constant tuning to achieve optimal performance. For SREs, managing fleets of VMs involves periodic tasks that can be daunting without proper tools.

The challenges include: 

  • Right Sizing and Capacity planning
  • Provisioning and Configurations
  • Hardening, Security and Compliance
  • Patch Management and Cost Optimization
  • Backup and Disaster Recovery
  • Automation and Orchestration
  • Monitoring and Diagnostics

Simplified Management of VMs on Cloud

Managing virtual machines (VMs) in the cloud makes it easier and more efficient for Site Reliability Engineers (SREs) and DevOps teams. Here are the key points highlighting this simplification:

  • Hardened OS Images: Cloud marketplaces offer readily available hardened operating system images. 
  • Simplified Remediation: Unlike traditional environments where writing complex scripts and routines was necessary for remediation, cloud platforms provide APIs. These APIs enable straightforward actions like:
  • Rebooting VMs
  • Resizing resources
  • Performing other necessary adjustments easily
  • Built-In Tools and Insights: Modern cloud platforms have various built-in tools and matrices that provide valuable insights. These tools help SREs make informed decisions.

Over time, the management of VMs has become less complex due to these advancements.

The Role of Platform Teams in the Autonomous Journey

Platform teams are essential for guiding organizations through their autonomous journey:

  • Establishing Standards: They enforce standards across VM fleets to ensure consistent performance.
  • Manned-Unmanned Teaming Concept: This approach combines human oversight with automated systems to enhance operational efficiency. It allows minimal human intervention while maximizing the use of automated assets.

Manned-Unmanned Teaming in Autonomous Operations

Manned-unmanned teaming is the operational deployment of manned and unmanned assets in concert with a shared mission objective. It is actually a battlefield terminology, wherein the priority is to have a minimal number of boots on the ground while we use unmanned assets as a force multiplier to ensure mission success with the highest efficiency and least casualties.

Key Steps for the Platform Team

Here, the platform team forms the core, which enforces a certain set of standards across the fleet so that the unmanned assets can identify applications, make sense of data, find anomalies, recommend fleet optimizations, and facilitate autonomous actions. The platform team sets the framework for autonomous systems to act upon. 

1. Identify App Boundaries

  • Group similar computing instances performing the same task, termed as an app.
  • Enable collective action if issues arise.

2. Metrics Standardization

  • In heterogeneous fleets, ensure correct labeling of metrics for accurate identification.
  • Different systems may use various exporters (e.g., Node Exporter for Linux, WMI Exporter for Windows).

3. Identifying the Golden Signal

Determine key telemetry signals to monitor, including:

  • Latency
  • Error rates
  • Saturation
  • Throughput

These signals help feed algorithms and machine learning systems for generating recommendations.

4. Application Discovery and Classification

  • Use traffic patterns and tagging to identify applications.
  • Group similar apps together.
  • If instances do not fit into these criteria, classify them as orphaned apps.

5. App to Metric Mapping

  • With the discovered app information, we can identify matrices and use them to run algorithms and find optimization opportunities and remediation. Establish a clear mapping between applications and their associated metrics.

The first five steps are relatively easier. Handling remediation is the last step and the toughest one.

6. Handling Remediations

The final step of remediation is the most challenging. If the recommendation system proposes a remediation action, it must follow a sequence of steps for safe execution in the customer environment.

Steps for Safe Remediation Action

Here are some steps for safe remediation action: 

  • Action Assessment: If a recommendation system proposes a remediation action, it must first determine if the action can be safely performed on the application without risk. If there is a green signal, the process continues.
  • Timing Consideration: Next, assessing whether it is the right time to apply the action or if there is a preferred time for executing it on the application is essential. If both the first two steps are approved, the action can proceed.
  • Action Execution: If both previous steps are clear, execute the action.
  • Health Check: After executing the action, the application's health needs to be evaluated. This step is crucial for ensuring that the application remains functional post-remediation.
  • Efficacy Validation: The final step involves validating the operation's efficacy. This is vital as it allows us to close the learning loop and utilize this information for future actions.

Sedai addresses the challenges SREs face using cloud and custom APIs to identify and discover infrastructure components. The inference engine employs this information to build a topology. With the topology, we can deduce the application. Our metric exporter collects data from all monitoring providers, and using the application and metrics data, Sedai’s machine learning algorithms generate remediations and optimization opportunities.

670d671d4b811ea5a544607d_AD_4nXfhkcuQoKfst93c0CwuhThJ73rbMBIj0gB3ppeXqv22QmRIaE1pcZAcmHcPiFCzj0iHkoX_XmFqmRsqyiBByXK2wW7qNbrDZYbuRumYbpRMjS1B5gWXCG3wTK-8H3hF5ij_5JzR2rxAhuRVnGuM84MfHVku.webp

This information is then handed off to the execution engine, which carries out the final leg of the process. The execution engine must be carefully and cohesively integrated into the platform to effectively utilize remediation APIs for executing actions in the cloud.

Autonomous Actions for VM Availability and Optimization

Autonomous actions for availability and optimization in virtual machines (VMs) involve executing autonomous remediations and also optimizing performance and cost.

  • Autonomous Remediations: Execute remediations based on anomalies detected in latency, errors, and custom metrics.
  • Performance Optimization: Recommend optimal auto-scaling configurations and vertical resizing for applications based on predicted saturation metric patterns.
  • Cost Optimization: Based on predicted traffic, identify the appropriate sizing required for applications. Use seasonal information to shut down development and QA instances selectively.

You can run autonomous remediations based on anomalies detected on latency error and custom matrices. When it comes to performance or cost optimization, you need to look at the provisioning state of the application. 

  • Over-Provisioned Applications: In cases of over-provisioning, costs can be saved by resizing the application.
  • Under-Provisioned Applications: If an application is under-provisioned, it is necessary to assess the performance impacts and resize it accordingly to remedy any problems.
  • Selective Shutdown: In certain scenarios, selective shutdown based on the seasonality of non-usage of the VMs can also be implemented.

Sedai VM Optimization 

670d671df41d4f15261a9acc_AD_4nXekiZ3jQWFKMrHHp01Q5sDKSlQT1QD6vbd-O7VnEwsqhMUl6_h9x4g9a1Dx9oESMovxJkEfKP1TKwt096GkTVEtl3VIabfdmeTxpm28VFaOa-RvpVFVsUR09lf0dNuJc9nQvq-furL2uyTqK1qPoQaoe18u.webp

In the Sedai application, the VM optimization page showcases optimization opportunities for specific accounts. It displays autonomous execution undertaken for VM-based applications, demonstrating meticulous attention to detail in each step.

670d671d9bf05f67088e8045_AD_4nXdeE6Ej2jkmgf7ZFPMLKOBf054NvmMfoDieftgAI94CIBqcu7Cf_kjpTxt6oCrljsN0alGpL5TR-qyrbbDcnnXsSCbvTmeWtSOnkKEYvQPt7gorThr5vkJfztZyhkpalj54BUUevEEpu3pMWVBTBh-fkWmt.webp

In environments with complex setups, Sedai provides autonomous systems with a higher degree of freedom. If desired, control can be handed over to a Site Reliability Engineer (SRE) for collaborative execution of autonomous actions. This aspect of teamwork in execution is referred to as manned and unmanned teaming.

Addressing non-standard setups in VM-based applications involves identifying multiple applications deployed on a single VM. This is done by examining the ports used and deriving metrics from the gathered information. While discovering application data is relatively easier, handling the operations can be more complex. 

The autonomous system ensures that the right time for executing operations is considered, particularly for multiple applications running on the same VM. It is safer to schedule operations during low-traffic periods or planned downtimes.

Support for VM-Based Applications in Clouds

Sedai does support VM-based applications across various cloud platforms. Currently, we provide support for:

  • AWS: Full support for virtual machines on Amazon Web Services.
  • Azure: Comprehensive support for VM-based applications on Microsoft Azure.
  • VMware: Support for enterprise customers using VMware.
  • On-Premise: Supported through custom integrations.

Sedai plans to broaden its product rollout for Google Cloud Platform (GCP), expected by the end of Q4. This multi-cloud support ensures users can leverage Sedai's capabilities regardless of their cloud environment.

Unlock the Future of Autonomous VM Management with Sedai

In cloud computing, Sedai stands at the forefront of optimizing VM management through its innovative autonomous systems. By simplifying the complexities of virtual machine operations, Sedai enables businesses to enhance performance, reduce costs, and ensure robust application availability. 

With support for major cloud platforms like AWS, Azure, and VMware, along with plans for GCP, Sedai provides a comprehensive solution tailored to meet the diverse needs of enterprises. 

Ready to optimize your VM management and elevate your cloud operations? Discover how Sedai can transform your infrastructure with autonomous solutions tailored to your business needs. 

Book a demo today to learn more about our offerings and get started on your journey to enhanced efficiency and effectiveness in cloud computing.