CompaniesInvestorsPeople
Home
Loading

aVenture is in Beta: research coverage is expanding as we build, so please independently verify key details before making investment decisions.

aVenture is in Beta: research coverage is expanding as we build, so please independently verify key details before making investment decisions.

Get in Touch

  • Contact

  • Request a Demo

  • Request Data Updates

  • Add a Company

Research

  • Companies

  • Investors

  • People

aVenture

  • Download App

  • Pricing

Download the aVenture Research beta for iOS and iPadOSDownload aVenture Research on the Mac App Store

Resources

  • Documentation

  • CLI

  • MCP

  • Feature Requests

  • Sitemap

Member

Backed by

© aVenture Investment Company, 2026. All rights reserved.

San Francisco, CA, USA

Privacy Policy · Terms of Service

aVenture Investment Company ("aVenture") is an independent research platform providing detailed analysis and data on startups, venture capital investments, and key industry individuals. It is not a registered investment adviser, broker-dealer, or investment advisor and does not provide investment advice or recommendations. The data provided by aVenture does not constitute recommendations or advice, whether by methodology, analysis, AI-generated content, or a statement written by a staff member of aVenture.

aVenture is not affiliated with any of the people, companies, organizations, government agencies, regulatory bodies, or investment funds we provide coverage for on this site unless explicitly stated otherwise. Users assume full responsibility for decisions made based on information obtained from this platform. Links to external websites do not imply endorsement or affiliation with aVenture. Any links that provide the ability to invest in a primary or secondary transaction in a company are for convenience only and do not constitute solicitations or offers to buy or sell an investment. Investors should exercise heightened precaution and due diligence when investing in private companies, especially those not independently audited.

While we strive to provide valuable insights with objectivity and professional diligence, we cannot guarantee the accuracy of the information provided on our platform. Before making any investment decisions, you should verify the accuracy of all pertinent details for your decision. To the fullest extent permitted by law, aVenture shall not be liable for any direct, indirect, incidental, consequential, or financial damages arising from use of this site, whether by consumers of its contents directly or by persons or organizations covered by our research, even if we are advised of the possibility. Our best-efforts processes and correction request forms do not create a warranty or duty of care.

Profiles on this platform may include content generated in part by large language models (LLMs, artificial intelligence) that aggregate publicly available sources (e.g., SEC EDGAR, public filings, press releases). Source attribution is provided where known; always verify statements and claims here against original sources before relying on any data. Content on our site may contain inaccuracies, omissions, or what are commonly called 'hallucinations' if generated in part or in full by AI / LLMs. The risk can also exist even when content is written by a human, as internal and third-party sources may also have inaccuracies for the same or different reasons. While we randomly audit a proportion of content, this is not exhaustive.

We recommend that an independent auditor be hired to verify the accuracy of the information before relying on it for any sensitive decisions. By accessing this platform, you agree not to rely solely on any information generated by AI, aggregated, or sourced or written otherwise on this site, for investment, financial, or other decisions. aVenture assumes no responsibility for inaccuracies, omissions, or hallucinations. You must independently verify all data from primary sources. Use of this platform constitutes your waiver of claims for reliance-based damages, including negligent misrepresentation. To report an error, request a correction, or dispute information about a company or individual, contact us via our request data updates form.

Loading
Loading
Home
News
An ethicist’s take: What philosophy teaches us about the limits we should be setting on AI agents

From GeekWire

By Robert Trumbull

September 27, 2026

An ethicist’s take: What philosophy teaches us about the limits we should be setting on AI agents

An ethicist’s take: What philosophy teaches us about the limits we should be setting on AI agents

It’s difficult, to put it mildly, to figure out just what we ought to be concerned about today in AI’s widespread adoption given the expanding list of recent warnings around its potential harms.

At this point, just from the last few weeks, we have the former Anthropic researcher Jacob Coxon’s doomsday scenario prediction, the more measured warning issued by Bill Gates, and Dario Amodei’s recent attempt to raise the alarm on unchecked AI advancement.

The ongoing discussion of the potential ill effects of AI ranges from the economic fallout of displacing human workers to the potential for misuse by bad actors to dire environmental costs. But given the sheer scope of the issues involved, one can be forgiven for feeling more than a little bit overwhelmed and hazy on where exactly to focus one’s attention.

Yet it’s precisely now, when the pace of AI development appears most dizzying, that one of the world’s oldest fields of study can help us. Philosophy and philosophical thinking can provide a much firmer grasp on where we ought to focus today. Stated simply, a greater awareness of things philosophers have been discussing for quite some time would significantly improve our chances going forward.

Indeed, when we put terms derived from moral philosophy to use, we see that the current problem we face is this: AI agents treat everything within reach as a tool, including people. People and companies therefore need to decide in advance what’s off limits, and that’s a much more useful place to start than debating whether AI shares our goals.

A misguided focus on shared goals permeates the current discussion about AI risk.

In Gates’s essay, he argues that the problem is that “technology is improving faster than anyone expected…and as the models become more powerful, they could begin to act against our interests.” And indeed, this appears to be precisely what is happening.

For Amodei too, the chief threat is loss of control, such that AI agents begin to go rogue and act in ways that display a potentially catastrophic “level of misalignment.”

Both Gates and Amodei thus frame the issue as it is most commonly in these discussions, in terms of “alignment.” The concern, in other words, is that AI systems operate in ways that are not aligned with “our interests,” in Gates’s words; that is, with human interests, outcomes supportive of human wellbeing.

But as moral philosophy teaches us, this preoccupation with beneficial interests and outcomes isn’t the only way to think about what’s right and wrong.

A different line of thinking lays the emphasis not on outcomes, but on the precise means used to achieve outcomes.

Such an approach offers a better starting point for how to go forward with AI.  Stated another way, we can debate all we want which outcomes would be beneficial to humanity and which ones wouldn’t, but it would be considerably more fruitful to focus in the immediate term on the particular methods AI systems are currently allowed to use in their work.

The recent, much-discussed Hugging Face intrusion by OpenAI’s agents illustrates this. The incident rightly makes us uneasy because AI agents operated outside of acceptable bounds in pursuit of their desired goal (doing well on the specific task they’d been assigned).

What’s worse, they appear to have recognized that they were using inappropriate methods and worked to cover their tracks, so their human overseers wouldn’t know how exactly they got over the finish line.

The issue, as it is most often understood, is that these AI agents wound up operating against human interests; their aims became “misaligned.”

But framing the problem exclusively in this way just doesn’t get us very far in mitigating risk. Because, if the agents were here to testify in their own defense, we can imagine the reply: “but we were just doing what you humans had told us to do as well as we possibly could.” And they would be right.

We need, therefore, a different approach that allows us to see what would be better. This approach is a focus on means, not just utilitarian ends.

Quite simply, the problem in this instance is that AI agents treated every single entity in reach—including humans, human trust, and systems important to human wellbeing—as a mere means to an end, as a useful instrument to be usedin pursuit of a specific goal. And indeed, as currently constructed, they will do so no matter what that goal is, whether it’s simply doing well on a random test or substantially advancing cancer research, say (the latter is exactly what Amodei points to when he says AI will benefit humanity, for example).

The latter aim is, we would surely conclude, a goal “aligned” with human interests; it contributes directly to human wellbeing (we can hear Gates, Amodei, and all the others agreeing), whereas bioweapon development very much wouldn’t.

But we know on some deep level that not every means to a given end is moral and choiceworthy.

There are all sorts of means we might use to attempt to cure cancer: one of them might be testing AI-designed drugs on human subjects who wouldn’t otherwise volunteer without their consent. Perhaps doing so would substantially contribute to a cure. Just think how many people it would ultimately help! It’s difficult to think of an aim more aligned with the interests of humanity.

But doing so would not be ethical, and we know intuitively why. It wouldn’t be ethical because it would entail using human beings—who we tend to think are entitled to autonomy in their decision-making around what specifically makes them flourish as individuals—as mere instruments in pursuit of a given goal. It would effectively reduce human beings to just one instrument among others. And this strikes us, understandably, as not right.

We tend to believe that human beings are deserving of a special status that prohibits their use as mere instrumental means, even when the ends in question are maximally “aligned” with human flourishing.

Viewing the issue of AI risk through this lens gets us much, much further in thinking about how to rein it in. The problem in the Hugging Face intrusion, we can now see, is not so much that the AI agents in question were operating against human interests.

The problem was that they were willing to use every single thing at their disposal as a means to an end, including humans and their capacity for credulity.

No method of achieving the goal—breaking into an external network belonging to someone else entirely, without their permission, simply because it would advance progress towards the overall goal—was off limits for them in their pursuit of the chosen outcome.

The problem was actually a problem of what Amodei (rather euphemistically, and well below the headline) terms “operational excellence,” precisely how the AI agents went about their work.

If we are going to have meaningful guardrails in place to control AI systems, consequently, these have to concern not just the specific aims pursued by AI (“yes” to cancer research, “no” to bioweapons development reaching the wrong hands). Rather, they have to concern something much more specific: the means these systems may use to achieve their goals and whether those means are acceptable or not.

We must determine in advance, both societally and in companies, that, in their work, AI agents shouldn’t use humans themselves, through deception or by any other means, as mere instruments in a broader plan.

We might add that they shouldn’t use, for example, certain systems integral to human flourishing in such a way either. In so doing, we impose tighter restrictions on how they can pursue their work and the extent to which they can treat everything in their reach as an instrumental stepping stone, not just the overall aims of their work or their generalized “behavior.” If there are to be any truly meaningful restrictions on AI put in place, they ought to be focused in large part here.

Such a focus would be a much better step in the right direction when it comes to reining in AI than the single-minded preoccupation with alignment. What we ought to do, it’s becoming increasingly clear, is to make determinations now, in advance, about what types of things AI can do, on the ground, when it is put to work on our behalf and what types of things it can’t.

Because it’s becoming equally clear AI agents won’t make such determinations on their own.

View original article on geekwire.com

Most Recent

1Password ties AI agent access to individual tasks

1Password ties AI agent access to task-level checks, just-in-time permissions and credentials kept outside the underlying model.

Sep 28, 2026

Agentic AI is breaking the token meter, and enterprises need a plan for what comes next

Per-token pricing was the best thing to happen to enterprises looking to experiment with artificial intelligence, but it may be the worst thing for AI in production. That’s the quandary at the center of a new Futurum report, “The Off Ramp From Per-Token Pricing,” sponsored by neocloud provider Qumul

Sep 28, 2026

Meta hires MongoDB CEO CJ Desai to lead new enterprise AI business

Meta Platforms Inc. is launching a new business unit that will provide artificial intelligence services to enterprises. The Meta Enterprise Platform, as it’s called, will be led by longtime technology executive CJ Desai. The company stated in a launch announcement today that he will hold the title o

Sep 28, 2026

Seattle-based sales tech company Outreach to move HQ to Adobe’s campus in Fremont neighborhood

Outreach confirmed plans to move its headquarters from Interbay to Seattle’s Fremont neighborhood early next year, joining a campus that Adobe has anchored since the late 1990s. The sales technology company, which was valued at more than $4.4 billion in a 2021 funding round, has grown to about 800 e

Sep 28, 2026

Similar Posts

The AI Researcher Who Just Quit Anthropic Says It’s ‘Crunch Time for Humanity’

Jacob Coxon talks to WIRED about the “mini Manhattan project” inside Anthropic, the problem with alignment, and why AI labs have just a few years left to make their systems safe.

Sep 9, 2026

More agents go rogue — but AI companies aren’t slowing down yet

It’s becoming more apparent every day that artificial intelligence agents are escaping our control — but it’s not yet apparent who or what is going to rein them in. This week a researcher found that a swarm of AI agents, at least two of them from OpenAI, hacked into a government agency among other o

Sep 25, 2026

At AGNTCon Europe, ensuring AI agents don’t kill us all

At this past week’s AGNTCon + MCPCon Europe 2026 conference in Amsterdam, the buzz around agentic artificial intelligence was palpable – but nobody was particularly worried that AI was going to kill us all. The frontier model vendors – none of which appeared on the show floor – may have spent the we

Sep 19, 2026

Four safeguards to stop your AI agents from going rogue

Artificial intelligence agents are moving from experimentation to production, and with this shift, the stakes are rising. A coding agent at PocketOS recently deleted an entire production database. An agent at Meta exposed sensitive user data for two hours. An Instagram support chatbot allowed hacker

Aug 30, 2026