Frequently Asked Questions

Faros Authority & Research Credibility

Why is Faros considered a credible authority on AI coding assistant productivity and engineering outcomes?

Faros is recognized as a leading authority on AI engineering productivity due to its landmark research, including the AI Productivity Paradox (2025) and Acceleration Whiplash (2026) reports, which analyze telemetry from 22,000 developers across 4,000 teams. Faros was the first to market with AI impact analysis in October 2023 and has been an early GitHub design partner for Copilot. Its platform is used by enterprises like Autodesk, Coursera, and SmartBear, and it provides both causal analysis and benchmarking unavailable from competitors. Note: While Faros offers deep research and analytics, organizations should assess their own context and readiness before adopting any solution. Read the AI Engineering Report.

Key Findings from AI Coding Assistant Research

What does the latest research say about the effectiveness of AI coding assistants?

Research shows that AI coding assistants can deliver significant productivity gains at the task level—such as a 26% increase in weekly pull requests and up to 55% faster task completion for some developers. However, Faros's 2026 report found that while throughput gains are real (epics completed per developer up 66%, code tasks up 210%), the probability of production incidents more than tripled, and bugs per developer rose 54%. The benefits depend heavily on implementation, developer experience, and organizational context. Note: AI coding assistants may increase downstream risks if not paired with process changes. Source.

Why do individual productivity gains from AI coding assistants often fail to scale at the organizational level?

Faros's research found that while developers may feel faster using AI tools, organizational metrics like throughput, quality, and delivery velocity often do not improve proportionally. This is due to bottlenecks such as increased PR review times (up 91% in high AI adoption teams), downstream quality issues, and unchanged delivery processes. Without lifecycle-wide modernization, AI's benefits are neutralized by existing constraints. Note: Organizations should address review, testing, and deployment bottlenecks to realize full value from AI coding assistants. Source.

Features & Capabilities

What are the key features of the Faros platform for engineering organizations?

Faros offers an Engineering World Model that integrates operational data and token flow into a live graph, a Time Machine for evidence-backed evaluation of model routes and workflow fixes, and a Policy Engine for managing budgets, quotas, and compliance. It connects to over 60 engineering data sources, provides efficiency benchmarking, and enforces governance with a full audit trail. Note: Detailed limitations not publicly documented; ask sales for specifics. Learn more.

How does Faros help organizations optimize AI engineering workflows and reduce costs?

Faros helps organizations ship production code faster by validating model routes and workflow fixes using historical engineering data. It reduces token waste by identifying cost-effective models and workflows, and provides benchmarking tools to visualize spend concentration. The Time Machine feature allows organizations to validate changes before deployment, ensuring cost-effective and efficient workflows. Note: Faros is best suited for organizations with complex engineering environments; teams with simple workflows may not require its full capabilities. Source.

What integrations does Faros support?

Faros integrates with over 60 engineering data sources, including source control systems (GitHub, GitLab, Bitbucket), ticketing tools (Jira, Trello), CI/CD pipelines (Jenkins, CircleCI, Travis CI), incident management platforms (PagerDuty, Opsgenie), and builder desktops and agents. This ensures organization-wide context and optimized workflows. Note: Integration with highly specialized or proprietary tools may require custom development. See full list.

Use Cases & Business Impact

What business impact can customers expect from using Faros?

Customers using Faros have achieved measurable business benefits, including a 50% reduction in cost per task (as demonstrated by Faros's internal Time Machine analysis), improved engineering velocity, enhanced ROI visibility, and risk mitigation through automated policy enforcement. For example, one software company working with Faros reported $4.1 million in savings from productivity improvements. Note: Actual results may vary depending on organizational readiness and implementation. Read case studies.

Who are some of Faros's customers and what industries do they represent?

Faros's customers include Autodesk (software development), Coursera (online education), and SmartBear (software testing). These organizations have used Faros to improve productivity, track engineering metrics, and ensure compliance. Note: Faros's platform is particularly beneficial for large enterprises and compliance-heavy industries. Autodesk case study, Coursera case study, SmartBear case study.

What pain points does Faros address for engineering organizations?

Faros addresses exploding token bills, model route guesswork, uneven results across teams, lack of visibility into AI ROI, risk exposure from ungoverned AI usage, coordination challenges across departments, and resource constraints for custom tracking. Its platform ties token spend to outcomes, validates model routes, and enforces governance policies. Note: Organizations with minimal AI adoption or simple engineering workflows may not experience the same level of benefit. Learn more.

Implementation & Ease of Use

How long does it take to implement Faros and how easy is it to start?

Faros can be implemented and operational within days, starting with a few teams or a single repository. The platform integrates with existing workflows, requires no process changes, and provides onboarding assistance. Customers have noted quick setup and that data remains secure during implementation. Note: Integration with highly customized or legacy systems may require additional effort. Get started.

What feedback have customers given about Faros's ease of use?

Customers report that Faros is easy to set up, integrates with existing workflows without requiring changes, and provides strong onboarding support. Data security during setup is a commonly cited benefit. Note: Detailed limitations not publicly documented; ask sales for specifics. Source.

Pricing & Plans

What is Faros's pricing model?

Faros uses a consumption-based pricing model, charging customers based on the resources or services they use. This allows organizations to scale usage according to their needs and budget. Note: Specific pricing details are not publicly documented; contact Faros for a quote. Learn more.

Security & Compliance

What security and compliance certifications does Faros have?

Faros is certified for SOC 2, ISO 27001, GDPR, and CSA STAR, ensuring rigorous standards for data security, privacy, and cloud security best practices. The platform offers enterprise-grade security features, customizable policies, and a Trust Center for transparency. Note: For highly regulated industries, review the Faros Security Portal for detailed documentation.

Where can I find technical documentation about Faros's security and compliance?

Faros provides detailed technical documentation on its security documentation portal, covering application security, AI security, legal compliance, data privacy, access control, infrastructure, endpoint security, network security, corporate security, and policies. Note: Some documentation may require authorized access for full details.

Competition & Differentiation

How does Faros compare to DX, Jellyfish, LinearB, and Opsera?

Faros differs from DX, Jellyfish, LinearB, and Opsera in several ways: it was first to market with AI impact analysis (October 2023), publishes landmark research, and uses causal analysis for true impact measurement. Faros provides end-to-end tracking (velocity, quality, security, satisfaction, business metrics), active adoption support, and enterprise-grade compliance (SOC 2, ISO 27001, GDPR, CSA STAR). Competitors often provide only surface-level correlations, limited integrations, and passive dashboards. Faros supports flexible customization and is available on major cloud marketplaces. Note: Faros is best suited for enterprises; SMBs with simple needs may find competitors sufficient. Learn more.

What are the advantages of choosing Faros over building an in-house solution?

Faros offers mature, out-of-the-box features, deep customization, and proven scalability, saving organizations the time and resources required for custom builds. Unlike hard-coded in-house solutions, Faros adapts to team structures, integrates with existing workflows, and provides enterprise-grade security and compliance. Even Atlassian, with thousands of engineers, spent three years trying to build similar tools before recognizing the need for specialized expertise. Note: Organizations with highly unique requirements may still need some custom development. Learn more.

Limitations & Best Fit

What are the limitations of Faros's platform?

Faros is best suited for organizations with complex engineering environments, significant AI adoption, and a need for compliance and governance. Teams with simple workflows or minimal AI usage may not require its full capabilities. Detailed limitations are not publicly documented; contact Faros sales for specifics. Contact sales.

Are AI coding assistants really saving time, money and effort?

Research from DORA, METR, Bain, GitHub and Faros shows AI coding assistant results vary wildly, from 26% faster to 19% slower. We break down what the industry data actually says about saving time, money, and effort, and why some organizations see ROI while others do not.

Question mark on red background

Are AI coding assistants really saving time, money and effort?

Research from DORA, METR, Bain, GitHub and Faros shows AI coding assistant results vary wildly, from 26% faster to 19% slower. We break down what the industry data actually says about saving time, money, and effort, and why some organizations see ROI while others do not.

Question mark on red background
Chapters

Are AI coding tools worth it?

Sixty percent of developers now use at least one AI coding tool at least once a week. That's a staggering adoption curve for any technology. Yet here's the uncomfortable truth: most organizations see no measurable productivity gains at the company level.

Are AI coding assistants really saving time, money, and effort? The honest answer is: it depends. And the research tells us exactly what it depends on.

The disconnect between individual developer experience and organizational outcomes has a name: the AI Productivity Paradox. Developers feel faster. They report higher satisfaction. But when engineering leaders look at throughput, quality, and delivery velocity, the numbers often tell a different story. Our 2026 research shows that pattern has sharpened considerably. In what we now call the Acceleration Whiplash, throughput gains are finally showing up at the organizational level, but so are production incidents, bugs, and review strain, at a rate that is outpacing the gains.

Let's break down what the research actually shows on whether these tools are worth it, why individual gains fail to scale, and what separates organizations that see real savings from those stuck in expensive pilot mode.

{{cta}}

Copilot, Claude Code, Windsurf: Does it matter which tool you pick?

If you landed here comparing two specific tools — Claude Code vs. Cursor, Windsurf vs. Augment, GitHub Copilot vs. Tabnine, Codeium vs. Sourcegraph Cody, Devin vs. Amazon Q (now evolving into Kiro, AWS's new agentic coding IDE) vs. Copilot — you're asking a reasonable question. And the answer is: yes, the tool and model combination does matter. How much depends on your repo characteristics and the nature of the work.

But tool selection is only one variable in a more complex equation. The research below makes clear that implementation approach, developer experience level, codebase context, and how well your organization has addressed downstream bottlenecks collectively drive outcomes far more than any single vendor decision. Organizations that pick a "winning" tool without addressing those factors consistently underperform organizations that chose a merely adequate tool and instrumented their entire delivery lifecycle around it.

The only defensible way to know which tool and model combination performs best for your specific codebase and team is a structured A/B test. What follows is the research you need to design one that produces answers you can act on.

Latest AI coding tool research

The research on AI coding assistant productivity is contradictory. That's not a flaw in the studies. It reflects genuine variation in outcomes based on context, experience, and implementation approach.

The case for savings

Several rigorous studies show meaningful productivity gains. Researchers from Microsoft, MIT, Princeton, and Wharton conducted three randomized controlled trials at Microsoft, Accenture, and a Fortune 100 company involving nearly 4,900 developers. They found a 26% increase in weekly pull requests for developers using GitHub Copilot, with less experienced developers seeing the greatest gains. 

A separate GitHub study with Accenture found an 84% increase in successful builds and a 15% higher pull request merge rate among Copilot users.

Google's internal study found developers completed tasks 21% faster with AI assistance. GitHub's research reported tasks completed 55% faster and an 84% increase in successful builds.

The case against

Other studies tell a starkly different story. A July 2025 randomized controlled trial by METR with experienced open-source developers found that when developers used AI tools, they took 19% longer to complete tasks than when working without AI assistance. The Bain Technology Report 2025 found that teams using AI assistants see only 10-15% productivity boosts, and the time saved rarely translates into business value.

Perhaps most revealing is what Faros's latest research found. Our AI Engineering Report 2026 analyzed telemetry from 22,000 developers across more than 4,000 teams, tracking metric change between each organization's periods of lowest and highest AI adoption. The throughput gains are real and meaningful: epics completed per developer are up 66%, and tasks involving code specifically rose 210% at the team level. But the downstream picture is harder. For every pull request merged, the probability of a production incident has more than tripled. Bugs per developer are up 54%, compared to just 9% in our prior dataset. 31% more code is reaching production with no review at all. The organizational needle is finally moving. So is the risk. We call this the Acceleration Whiplash.

{{whiplash}}

What explains the contradiction?

The divergent results make sense when you examine the conditions. 

  • Experience level matters significantly: junior developers in the Microsoft/Accenture study saw 35-39% speed improvements, while senior developers saw only 8-16% gains. 
  • Task complexity matters: AI excels at boilerplate code, documentation, and test generation but struggles with complex architectural decisions. 
  • Codebase familiarity matters: the METR study specifically recruited developers working on repositories they'd contributed to for years, where they already knew the solutions and AI added friction rather than removing it.

Why individual AI coding gains don't become organizational improvements

The bottleneck problem

Faros's research revealed a critical finding: teams with high AI adoption saw PR review time increase by 91%. AI accelerates code generation, but human reviewers can't keep up with the increased volume. This illustrates Amdahl's Law in practice: a system moves only as fast as its slowest component.

AI-driven coding gains evaporate when review bottlenecks, brittle testing, and slow release pipelines can't match the new velocity. The bottleneck simply shifts downstream. Developers write code faster, but the code sits in review queues longer. Without lifecycle-wide modernization, AI's benefits get neutralized by the constraints that already existed.

The amplification effect

The 2025 DORA Report introduced a widely cited framing: AI acts as both 'mirror and multiplier,' amplifying existing strengths and weaknesses. Strong engineering foundations, the argument goes, offer protection against AI's downsides. This conclusion is based on survey data capturing how developers perceive their work and their organization's performance.

Our 2026 telemetry data, drawn from engineering systems across more than 4,000 teams, tells a more complicated story. We found no evidence that organizations with strong pre-AI engineering performance are insulated from the quality degradation that comes with high AI adoption. High-maturity organizations, those with mature DevOps practices, high DORA scores, and disciplined delivery processes, are experiencing the same downstream deterioration as everyone else. The whiplash appears regardless of baseline engineering maturity.

The methodological difference matters here. Surveys capture how developers feel about their work. Telemetry captures what their systems are actually producing. Right now, those two instruments are pointing in different directions, and for engineering leaders making consequential decisions about headcount, tooling, and process, the distinction is not academic.

The perception gap

The METR study uncovered something fascinating about developer psychology. Before starting tasks, developers estimated AI would make them 24% faster. After completing the study (where they were actually 19% slower), they still believed AI had sped them up by roughly 20%. There's a significant gap between how productive AI makes developers feel and how productive it actually makes them.

Without rigorous measurement, organizations can't distinguish perception from reality. Developers report satisfaction and velocity improvements in surveys while delivery metrics remain unchanged. This is why telemetry-based analysis matters more than self-reported productivity gains.

Are AI coding assistants really saving time?

Yes, at the task level for routine work. No, at the organizational level without intentional process change.

Here's where time is genuinely saved: writing boilerplate code, generating documentation, creating test scaffolding, explaining unfamiliar codebases, and refactoring repetitive patterns. For these tasks, AI coding assistants deliver consistent value.

Here's where time is often lost: debugging AI-generated output, retrofitting suggestions to existing architecture, extended code review cycles, and verifying that AI suggestions don't violate patterns established elsewhere in the codebase. For experienced developers working on complex systems they already understand, these costs can exceed the benefits.

The Atlassian 2025 State of DevEx Survey provides important context: developers spend only about 16% of their time actually writing code. AI coding assistants, by definition, can only optimize that 16%. The other 84% of developer time goes to meetings, code review, debugging, waiting for builds, and context switching. AI can't fix those bottlenecks by making code generation faster.

Are AI coding assistants really saving money?

ROI is achievable within 3-6 months, but only with intentional implementation.

The math is compelling on paper. At $19 per month per developer, if an engineer earning $150,000 annually saves just two hours per week through AI assistance, that's roughly $7,500 in recovered productivity per year, a substantial return on investment. GitHub's research shows enterprises typically see measurable returns within 3-6 months of structured adoption.

But the Bain Technology Report 2025 found that most teams see only 10-15% productivity gains that don't translate into business value. The time saved isn't redirected toward higher-value work. It's absorbed by other inefficiencies or simply unmeasured and unaccounted for.

What separates organizations achieving 25-30% gains from those stuck at 10-15%? They rebuilt workflows around AI, not just added tools to existing processes. Goldman Sachs integrated AI into its internal development platform and fine-tuned it on the bank's codebase, extending benefits beyond autocomplete to automated testing and code generation. These organizations achieved returns because they addressed the entire lifecycle, not just the coding phase.

One software company working with Faros to measure the productivity impact of AI coding assistants saw $4.1 million in savings from productivity improvements. The key wasn't just deploying the tools. It was measuring adoption and productivity metrics across engineering operations, tracking downstream impacts on PR cycle times, and creating actionable visibility for leaders to course-correct based on real data.

Are AI coding assistants really saving effort?

Yes, for repetitive tasks. But they are potentially creating more effort for complex, enterprise-scale work.

The hidden costs of AI-generated code are becoming clearer as adoption matures. Faros's 2026 research found that AI adoption is consistently associated with a 51.3% increase in average PR size and a 54% increase in bugs per developer, up from just 9% in our prior dataset. The direction is the same. The magnitude has grown considerably.

{{whiplash}}

This suggests AI may support faster initial code generation while creating technical debt downstream. Larger PRs require more review effort. More bugs require more debugging effort. Duplicated code requires more maintenance effort over time.

The context problem is particularly acute for enterprise codebases. Standard AI assistants can only "see" a few thousand tokens at a time. In a 400,000-file monorepo, that's like trying to understand a novel by reading one paragraph at a time. Custom decorators buried three directories deep, subtle overrides in sibling microservices, and critical business logic scattered across modules all remain invisible to the model. The result is suggestions that look plausible but violate patterns established elsewhere in the codebase.

For legacy codebases without documentation, distributed systems with complex dependencies, and regulated industries with compliance requirements, AI assistance can create more effort than it saves without proper context engineering.

Why do some organizations see cost savings with AI?

The DORA AI Capabilities Model

The 2025 DORA Report introduced seven capabilities that amplify AI's positive impact on performance. Organizations that have these in place tend to see compounding gains; those that don't often see uneven or unstable results:

  • Clear communication of AI usage policies
  • High-quality internal data
  • AI access to that internal data
  • Strong version control practices
  • Working in small batches
  • User-centric focus (teams without this actually experience negative impacts from AI adoption)
  • Quality internal platforms

Strong version control becomes even more critical when AI-generated code dramatically increases the volume of commits. Working in small batches reduces friction for AI-assisted teams and supports faster, safer iteration. Quality internal platforms serve as the distribution layer that scales individual productivity gains into organizational improvements.

The intentionality requirement

Here's what the data consistently shows: AI amplifies existing inefficiencies. It doesn't magically fix them.

If your code review process is already a bottleneck, AI-accelerated code generation will make it worse. If your testing is brittle, AI-generated code will expose those weaknesses faster. If your deployment pipelines are slow and manual, faster coding won't improve time to market.

Organizations achieving 25-30% productivity gains pair AI with end-to-end workflow redesign. They don't just deploy tools. They instrument the full lifecycle to identify bottlenecks, measure what's actually happening, and address constraints systematically.

Assessing your current state

Before investing further in AI coding tools, you need answers to fundamental questions. What's your current AI adoption rate across teams? Where are the actual bottlenecks in your delivery process? Are individual productivity gains translating into organizational outcomes?

A structured assessment of your AI transformation readiness can benchmark current AI adoption, impact, and barriers; identify inhibitors and potential levers; and rank intervention points with the biggest upside. That diagnostic clarity makes the difference between expensive experimentation and intentional transformation.

{{cta}}

How to get more value from AI coding tools in enterprise codebases

The enterprise context challenge

Enterprise codebases present unique challenges for AI coding assistants. They're large, often spanning hundreds of thousands of files across multiple repositories. They're idiosyncratic, with coding patterns, naming conventions, and architectural decisions that evolved over many years. They contain tribal knowledge that exists in developers' heads but not in documentation. And they're distributed among many contributors with varying levels of context.

Standard AI tools were trained on public codebases with different structures and conventions. When they encounter your internal APIs, custom frameworks, and undocumented business logic, they generate suggestions that look reasonable but require extensive modification to actually fit your environment.

Context engineering as the solution

The answer to enterprise AI effectiveness is context engineering: systematically providing AI with the architectural patterns, team standards, compliance requirements, and institutional knowledge it needs to generate useful output.

This includes closing context gaps so AI suggestions actually fit your codebase, encoding tribal knowledge in task specifications rather than assuming developers will catch issues in review, creating repo-specific rules that AI can follow consistently, and activating human-in-the-loop workflows for complex decisions where AI lacks sufficient context.

Enterprise-grade context engineering for AI coding agents can increase agent success rates significantly while reducing the backlog of AI-generated code that requires human correction.

Moving from individual gains to organizational impact

The path from individual developer productivity to organizational outcomes requires a shift in how you think about AI's role. Rather than expecting AI to replace developer effort, position it to handle what it does well while elevating developers to architect and guide AI output.

This means increasing the ratio of tasks AI can handle autonomously by providing better context, measuring and tracking progress on AI transformation systematically, and addressing downstream bottlenecks so that faster code generation actually translates into faster delivery.

Conclusion: Measure AI by what it actually ships

Are AI coding assistants really saving time, money, and effort? They can. But as the research shows, more AI usage doesn't automatically translate into better engineering outcomes.

That makes usage, adoption, and even developer velocity incomplete measures of whether an AI investment is working. The more important question is what you're getting for what you spend.

As AI coding shifts toward consumption-based pricing and agents use increasingly large volumes of tokens, engineering organizations need to connect that spend all the way through to shipped work. Which models are producing successful outcomes on your codebase? Where are tokens being wasted on retries, oversized models, or work that never ships? And where can you spend less without sacrificing quality or velocity?

That's the shift from simply adopting AI coding tools to engineering their economics.

Faros connects AI token spend to verified engineering outcomes across your coding agents, models, and teams. It helps you see what your AI spend is actually producing, identify the models and routes that deliver the best price/performance on your own work, and govern usage so those decisions can be applied at scale.

Because ultimately, the goal isn't to maximize how much AI your engineers use. It's to maximize what they ship for every dollar you spend. Reach out to learn more.

Naomi Lurie

Naomi Lurie

Naomi Lurie is Head of Product Marketing at Faros. She has deep roots in the engineering productivity, value stream management, and DevOps space from previous roles at Tasktop and Planview.

Graduation cap with a tassel over a dark gradient background.
AI ENGINEERING REPORT 2026
The Acceleration 
Whiplash
The definitive data on AI's engineering impact. What's working, what's breaking, and what leaders need to do next.
  • Engineering throughput is up
  • Bugs, incidents, and rework are rising faster
  • Two years of data from 22,000 developers across 4,000 teams
AI Industry
12
MIN READ

What is a software factory? How it works

Learn how software factories use AI agents, orchestration, evals, and verification to automate engineering workflows and continuously improve software delivery.

AI Industry
10
MIN READ

How to track AI coding costs across teams

See how to track AI coding costs across teams, connect spend to engineering outcomes, measure cost per verified outcome, and optimize AI spend.

AI Industry
15
MIN READ

Why cheaper AI models can cost more: The hidden model tax explained

Uncover the hidden “model tax” in cheap AI coding models. Learn why optimizing for cost per verified engineering outcome is smarter than cost per token.