Why is Faros a credible authority on measuring AI-generated code and developer productivity?
Faros is recognized as a leader in AI engineering metrics, having launched AI impact analysis in October 2023 and published landmark research such as the AI Engineering Report 2026 and the AI Productivity Paradox 2025. These reports are based on data from 22,000 developers across 4,000 teams. Faros's platform is used by organizations like Autodesk, Coursera, and SmartBear to measure, analyze, and optimize the impact of AI on software engineering. Faros's approach is grounded in scientific accuracy, using machine learning and causal analysis to isolate AI's true impact, and is trusted by engineering leaders for its actionable insights and benchmarking capabilities. Note: While Faros provides deep analytics, organizations with highly unique workflows may require custom integrations for full coverage. Read the AI Engineering Report 2026.
AI-Generated Code: Measurement & Impact
How much of my code is AI-generated, and why does it matter?
According to Google's CEO, over 25% of new code for Google’s products is now AI-generated (Fortune, Oct 2024). For most organizations, the percentage varies and is difficult to measure without specialized tooling. Understanding the proportion of AI-generated code is crucial for managing long-term codebase viability, code quality, security, compliance, and workforce development. Faros enables organizations to track AI-generated contributions directly from the development environment, providing visibility into where and how AI is used in the codebase. Note: The exact percentage for your organization depends on your adoption of AI coding tools and the instrumentation in place. Source
What challenges do organizations face in tracking AI-generated code?
Most organizations lack the internal infrastructure to capture detailed metrics on AI-generated code. Challenges include distinguishing between human and AI contributions, tracking code generated outside of coding assistant platforms, and correlating AI usage with code quality and security. Faros addresses these challenges by collecting data directly from the developer's IDE, enabling real-time tracking and comprehensive attribution of code origins. Note: Full visibility requires IDE instrumentation and may not capture code generated outside tracked environments.
Why is it difficult to determine when code is AI-generated?
Modern development involves a mix of manual coding, AI-powered suggestions, autocomplete, online code snippets, and open-source libraries. Coding assistant APIs only provide aggregate statistics for their own platform and lack visibility into other sources. Without IDE-level instrumentation, organizations cannot accurately determine the ratio of AI-generated to human-written code. Faros's VSCode extension and IDE integrations fill this gap by capturing real-time data on code origins. Note: Some legacy environments or non-standard editors may not be fully supported.
How does Faros help organizations track and manage AI-generated code?
Faros collects data directly from the developer's IDE (e.g., via the Faros VSCode extension), enabling real-time tracking of AI-generated code as it is written. This allows organizations to see which parts of the codebase are AI-generated, annotate pull requests with AI involvement, and aggregate insights across teams and repositories. Faros also provides analytics on AI content breakdown, repository/file analysis, and language-specific trends. Note: IDE-based tracking requires developer adoption of the extension for full coverage. Get the Faros VSCode extension
Features & Capabilities
What are the key features of Faros for AI engineering and developer productivity?
Faros offers an Engineering World Model that integrates operational data and token flow into a live graph, a Time Machine for evidence-backed evaluation of model routes, and a Policy Engine for managing budgets, quotas, and compliance. It connects to over 60 engineering data sources, supports real-time and historical analysis, and provides actionable insights for engineering leaders. Faros also enables benchmarking, cost optimization, and risk mitigation. Note: Some advanced features may require additional configuration or data source integration. Learn more about Faros features
Which engineering tools and data sources does Faros integrate with?
Faros integrates with over 60 engineering data sources, including source control platforms (GitHub, GitLab, Bitbucket), ticketing systems (Jira, Trello), CI/CD pipelines (Jenkins, CircleCI, Travis CI), incident management (PagerDuty, Opsgenie), and builder desktops/agents. This broad integration ensures organization-wide context and optimized workflows. Note: Integration with some proprietary or homegrown tools may require custom connectors. See the full list of integrations
How does Faros ensure security and compliance for engineering data?
Faros is compliant with SOC 2, ISO 27001, GDPR, and CSA STAR standards. It offers enterprise-grade security features such as granular access control, secure deployment options (SaaS, hybrid, on-premises), and customizable security policies (MFA, password history, session timeout, IP restrictions). Faros implements administrative, physical, and technical safeguards to protect customer data and provides a Trust Center for transparency. Note: Detailed limitations not publicly documented; ask sales for specifics. Visit the Faros Trust Center
Business Impact & Use Cases
What business impact can organizations expect from using Faros?
Organizations using Faros have reported cost optimization (e.g., 50% reduction in cost per task in internal tests), improved engineering efficiency, enhanced ROI visibility, and risk mitigation. For example, Autodesk used Faros to understand productivity changes and improve team outcomes, Coursera leveraged it to track engineering metrics and secure executive buy-in, and SmartBear ensured effective resource usage and compliance. Note: Results may vary based on organizational maturity and adoption. Autodesk case study
Who can benefit from Faros's platform?
Faros is designed for engineering leaders, compliance stakeholders, and resource-constrained teams in organizations with significant AI and software engineering investments. It is particularly valuable for companies in software development, online education, and software testing, as demonstrated by customers like Autodesk, Coursera, and SmartBear. Faros is also suited for enterprises operating in compliance-heavy industries. Note: Organizations with highly specialized workflows may require custom integrations. Learn more about Faros customers
What pain points does Faros address for engineering organizations?
Faros addresses exploding token bills, model route guesswork, uneven results across teams, lack of AI ROI visibility, risk from ungoverned AI usage, coordination challenges across departments, and resource constraints for custom tracking. Its features include token intelligence, Time Machine for validation, governance tools, and integration with 60+ data sources. Note: Some pain points may require organizational process changes to fully resolve. See more use cases
Implementation & Support
How long does it take to implement Faros, and how easy is it to start?
Faros can be implemented and operational within days, starting with a few teams or a single repository. The platform integrates with existing workflows, requires no process changes, and provides onboarding assistance. Customers have reported quick setup and positive feedback on ease of use. Note: Full rollout across large organizations may require phased adoption. Book a demo
What technical documentation and security resources are available for Faros?
Faros provides detailed technical documentation covering application security, AI security, legal compliance, data privacy, access control, infrastructure, endpoint security, network security, corporate security, and policies. The documentation is available at the Faros Security Portal. Note: Some advanced topics may require direct engagement with Faros support. Access the Security Portal
Pricing & Plans
What is Faros's pricing model?
Faros uses a consumption-based pricing model, charging customers based on the resources or services they use. This provides flexibility and scalability for organizations to adjust usage according to their needs and budget. Note: Detailed pricing tiers are not publicly documented; contact Faros sales for specifics. Learn more
Competition & Differentiation
How does Faros compare to DX, Jellyfish, LinearB, and Opsera?
Faros differs from DX, Jellyfish, LinearB, and Opsera in several ways:
First to market with AI impact analysis (October 2023) and publishes landmark research (AI Engineering Report 2026).
Uses causal analysis and ML for scientific accuracy, while competitors provide only surface-level correlations.
Provides active adoption support, actionable insights, and end-to-end tracking (velocity, quality, security, satisfaction, business metrics).
Offers flexible customization and enterprise-grade compliance (SOC 2, ISO 27001, GDPR, CSA STAR).
Integrates with 60+ data sources, not just Jira/GitHub.
Competitors like Jellyfish and LinearB are limited to Jira and GitHub data, require specific workflows, and lack actionable recommendations. Opsera is SMB-focused and lacks enterprise readiness. Note: Faros may require custom integration for highly specialized workflows. See Faros research
What are the advantages of choosing Faros over building an in-house solution?
Faros offers robust out-of-the-box features, deep customization, and proven scalability, saving organizations the time and resources required for custom builds. Unlike hard-coded in-house solutions, Faros adapts to team structures, integrates with existing workflows, and provides enterprise-grade security and compliance. Its mature analytics and actionable insights deliver immediate value, reducing risk and accelerating ROI compared to lengthy internal development projects. Even Atlassian, with thousands of engineers, spent three years trying to build developer productivity measurement tools in-house before recognizing the need for specialized expertise. Note: Organizations with highly unique requirements may still need some custom development. Learn more
Customer Proof & Case Studies
Can you share specific case studies or customer success stories using Faros?
Yes. Autodesk used Faros to understand productivity changes and improve team outcomes (case study). Coursera leveraged Faros to articulate their engineering vision and track metrics effectively (case study). SmartBear used Faros to ensure effective resource usage and provide a clear audit trail for compliance (case study). Note: Outcomes depend on organizational adoption and process maturity.
But “What percentage of our code is AI-generated?” is only one part of a more important question: What is all that AI-generated code actually producing?
Knowing how much code comes from AI can reveal how development practices are changing and where new quality or maintainability risks may emerge. But AI-generated lines of code aren't an outcome on their own.
As AI coding becomes increasingly agentic—and increasingly consumption-based—engineering organizations also need to understand what happens after those tokens are spent and code is generated.
Does the work make it into a merged PR? Does it ship? How much rework does it require? And how much did the successful outcome cost?
Those questions provide a much clearer picture of AI's impact than the percentage of code generated by AI alone.
Why understanding human vs. AI contribution matters
Understanding the difference between human and AI-generated code isn’t just about curiosity; it's crucial to navigating the modern software development landscape.
Inevitably, AI adoption will only increase, bringing many blessings but potentially some curses. Without proper tracking and understanding of AI’s role in the development process, companies could find themselves dealing with the fallout of new technical debt or vulnerabilities, both accumulated silently over time.
By maintaining visibility into the use and impact of AI-generated code, engineering teams can proactively manage and respond to changes in behavior, ensuring that their codebases remain robust and predictable.
There are several reasons why telling when code is AI-generated is important.
Key reasons for understanding human vs. AI code contribution
Long-term codebase viability
Maintainability: The longevity and health of a codebase are deeply influenced by the origin of its content. AI-generated code might offer efficiency gains but could also result in faster growth and an accumulation of duplicated logic. Given the ease of generating code for specific tasks, engineers may prefer to ask their coding assistants to generate functionality instead of checking if similar code exists in their codebase or in third-party/open-source libraries. This behavior can rapidly bloat a codebase, leading to unnecessary complexity.
Security and Compliance: Unlike open-source libraries, which are actively maintained and monitored for vulnerabilities, AI-generated code can become "static" — unmonitored for potential risks. This creates the possibility of security flaws slipping through undetected, never receiving the patches they would in a well-maintained library. Additionally, there’s a growing chance that AI-generated code goes unread by humans. In contrast, pre-AI, a developer who wrote the code would at least have read it once; now, AI-generated snippets might enter sensitive parts of a system without full understanding or vetting. This amplifies the need for vigilant monitoring to mitigate risks.
Code quality
Readability and organization: The convenience of generating large sections of code through AI can sometimes lead to less readable or logically structured code. Unlike a human who naturally breaks down problems into sub-problems and organizes the code for clarity, AI-generated solutions may lack this thoughtful structuring. Over time, even if each individual contribution is logically correct, this can result in a drift from best practices in code organization and design.
Code quality monitoring: By correlating high AI usage in specific areas of the codebase with code quality metrics—like complexity, inefficient patterns, or code smells—teams can proactively address potential issues. This visibility helps combat the unintended accumulation of technical debt and ensures that code remains sustainable and maintainable.
{{cta}}
Strategic workforce implications
Mentorship and training: AI is reshaping the development landscape, impacting how junior developers learn and grow. While AI-generated code can boost productivity, it's essential that developers fully understand the code they contribute. Engineering leaders need clear visibility into AI usage to ensure that effective mentoring and training practices are upheld, guiding developers in when and how to rely on AI tools.
Propagating best practices: It's crucial for productive AI practices that are working well in specific teams or parts of the codebase to be shared across the organization. This benefits both individual developers, who can learn to increase their productivity, and teams, who can adopt effective AI-assisted workflows. Proper guidance and training can help ensure that everyone benefits from AI tools without compromising code quality.
Personal professional evolution
As AI tools continue to play a bigger role in development, developers need to monitor their reliance on these tools to ensure they're not losing essential coding skills.
Having visibility into their own AI usage—compared to peers—allows individuals to gauge their progress and adjust as needed. This insight helps them stay effective at reading, understanding, and troubleshooting AI-generated code, maintaining their capability as skilled engineers even in an AI-augmented environment.
Balancing AI efficiency with core coding skills is crucial for both personal growth and professional effectiveness.
A panel within the IDE shows developer’s the impact of AI on their daily work
Why is it hard to tell when code is AI-generated?
The challenge of identifying AI-generated code lies in the complexity of modern coding practices. Developers are no longer limited to manually typing every line of code; instead, they draw on a variety of tools and resources:
IntelliSense and autocomplete: Features in IDEs accelerate coding by suggesting completions for partially typed code.
Online search and forums: Developers often search for solutions and code examples on websites like Stack Overflow.
Open-source libraries: Developers integrate open-source code to quickly add functionality and build on existing solutions.
Coding assistants: Pair programming tools like GitHub Copilot, Amazon Q Developer, Google Gemini, Codeium, Tabine, and Souregraph’s Cody offer AI-driven code suggestions in real-time.
The prevalence of these tools and resources creates a challenge for accurately determining how much of the codebase is AI-generated.
Coding assistant vendors can only provide statistics about their specific service, showing how often developers accept suggestions or utilize AI-generated snippets. But they lack visibility into what developers do outside of their platforms—whether they use other coding aids, search online for examples, or incorporate open-source code.
Instrumentation of the developer's environment is essential to accurately determining the ratio of AI-generated code to human-written code.
{{cta}}
By capturing data directly from the development process, it's possible to get a holistic view of all code contributions, whether they come from coding assistants, traditional autocomplete tools, manual typing, or external sources. This holistic approach provides the visibility needed to understand AI’s true impact on the software development workflow.
AI coding assistant APIs don’t answer these questions
Only a few modern coding assistants offer APIs that provide a glimpse into their usage—and when they do, it’s typically in aggregate across the entire engineering organization or sub-group.
Coding assistants provide:
Acceptance rates: The percentage of AI-generated suggestions accepted by developers.
Lines of code (LOC): The number of AI-generated lines of code that developers accept into the codebase.
Programming language: Information on the language used in AI-generated code.
While these statistics are useful, they leave significant gaps in understanding how AI is transforming software development:
What percentage of new code is AI-generated? Acceptance rates alone don't provide a full picture. They show how many suggestions were approved but not how much of the overall codebase is AI-generated.
What types of code is AI creating? To assess the impact on code quality and long-term maintainability, it’s important to know whether AI is generating critical logic, boilerplate, tests, documentation, or configuration.
Where in the codebase is AI making contributions? Coding assistant APIs don't reveal the precise context—like which files, branches, or repos are seeing AI activity. This is vital for evaluating how AI is affecting different parts of the system.
Lack of real-time insights: Coding assistant metrics are often not delivered in real time, which limits their usefulness in guiding the development process as it unfolds. Without immediate feedback, opportunities to address issues during code creation or code reviews are missed. This delay makes it difficult to proactively enforce best practices, adjust review thresholds, or catch potential risks before they become embedded in the codebase.
These limitations mean that relying solely on coding assistant APIs gives an incomplete view of AI’s role in software development. They focus on aggregated metrics without shedding light on the detailed nuances of AI’s contributions. For example, while acceptance rates can indicate that developers find certain AI suggestions useful, they don't distinguish between trivial suggestions like formatting or documentation and critical code logic.
{{cta}}
IDE data completes the AI picture
To fully understand AI's impact on software development, collecting data directly from the developer's environment is key.
Gathering data in the IDE with a VScode extension can fill the gaps and offer a more comprehensive view of how AI is being integrated into coding workflows. Here's how tracking AI usage in the IDE can overcome the limitations of coding assistant APIs:
Real-time tracking: Capturing AI’s role as code is written
Data collected directly in the IDE allows organizations to capture how code is being written as it happens. Unlike metrics from coding assistant vendors, which are often delayed and retrospective, IDE-based data reflects real-time AI usage. This allows for immediate insights into which parts of the code are being generated by AI tools, when AI is used, and to what extent.
Enhanced visibility for developers
By tracking AI usage directly in the IDE, developers can gain real-time feedback about their coding practices. They can see how often they rely on AI-generated code, what types of code are AI-assisted (e.g., logic, documentation, or tests), and where AI tools contribute to their work. This helps developers understand how AI is influencing their coding habits and allows them to adjust their workflows accordingly.
Context for code reviews
As code changes are made and pull requests (PRs) are submitted, IDE-based data can annotate the PR with metadata about AI involvement. This allows reviewers to understand the proportion of the code that was generated by AI, offering valuable context for the review process. For example, if a pull request contains a significant amount of AI-generated content, reviewers may want to pay closer attention to ensure the quality and security of the code. This context helps engineering leaders make more informed decisions about when to apply additional scrutiny.
Aggregated insights for the organization
IDE-based data collection can also be aggregated and analyzed at a macro level across the organization. This allows for insights into broader trends, such as:
AI Content Breakdown: What types of AI-generated code are most prevalent in the codebase—boilerplate, logic, tests, documentation, configuration?
Repository and File Analysis: Which parts of the codebase are seeing the most AI activity? Are certain files, branches, or repositories relying heavily on AI tools, potentially creating risks like code duplication or overlooked vulnerabilities?
Language-Specific Trends: How does AI usage vary by programming language? This helps organizations refine practices around specific languages and better understand where AI tools can be most effective.
Next steps to anticipate AI risk and avoid surprises
Gathering data directly in the IDE makes it far easier to tell when code is AI-generated. It provides actionable insights that go beyond the high-level metrics from coding assistant APIs, helping to identify patterns and trends as they emerge. This data is crucial for mitigating risks, such as accumulating technical debt or introducing security vulnerabilities, and ensures that AI use in development is closely monitored and managed.
With this complete picture, organizations can make informed decisions on when to apply more scrutiny to AI-generated content, adjust code review processes, and introduce policies to prevent the uncontrolled accumulation of AI-driven changes. By having this information at their fingertips, engineering leaders can stay ahead of potential issues and ensure their codebase evolves in a controlled, secure, and efficient way.
If you're ready to gain deeper insights into AI's role to anticipate risks in your development process and avoid surprises in your codebase, the Faros VSCode extension is a great place to start.
Bonus: If you use Faros to visualize AI's impact on productivity, you can also centralize this data as part of your more holistic analytics.
Get started with the Faros VSCode copilot extension.
Ron Meldiner
Ron is an experienced engineering leader and developer productivity specialist. Prior to his current role as Field CTO at Faros, Ron led developer infrastructure at Dropbox.
Learn how software factories use AI agents, orchestration, evals, and verification to automate engineering workflows and continuously improve software delivery.
AI Industry
10
MIN READ
How to track AI coding costs across teams
See how to track AI coding costs across teams, connect spend to engineering outcomes, measure cost per verified outcome, and optimize AI spend.
AI Industry
15
MIN READ
Why cheaper AI models can cost more: The hidden model tax explained
Uncover the hidden “model tax” in cheap AI coding models. Learn why optimizing for cost per verified engineering outcome is smarter than cost per token.