Post

Log inSign up

Post

Log inSign up

Anthropic on X: "Even while recording these demos, we encountered some amusing moments. In one, Claude accidentally stopped a long-running screen recording, causing all footage to be lost. Later, Claude took a break from our coding demo and began to peruse photos of Yellowstone National Park."

@AnthropicAI
Anthropic
@AnthropicAI
Even while recording these demos, we encountered some amusing moments. In one, Claude accidentally stopped a long-running screen recording, causing all footage to be lost. Later, Claude took a break from our coding demo and began to peruse photos of Yellowstone National Park.
00:00
3:06 PM · Oct 22, 2024·
530K
Views
122

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email

Relevant people

Avatar
Anthropic@AnthropicAIFollow
We're an AI safety and research company that builds reliable, interpretable, and steerable AI systems. Talk to our AI assistant @claudeai on https://t.co/FhDI3KQh0n.

Trending now

Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • @AnthropicAI
    Anthropic
    @AnthropicAI
    Oct 22, 2024
    Introducing an upgraded Claude 3.5 Sonnet, and a new model, Claude 3.5 Haiku. We’re also introducing a new capability in beta: computer use. Developers can now direct Claude to use computers the way people do—by looking at a screen, moving a cursor, clicking, and typing text.
    A benchmark comparison table showing performance metrics for multiple AI models including Claude 3.5 Sonnet (new), Claude 3.5 Haiku, GPT-4o, and Gemini models across different tasks.
    462
  • @AnthropicAI
    Anthropic
    @AnthropicAI
    Oct 22, 2024
    The new Claude 3.5 Sonnet is the first frontier AI model to offer computer use in public beta. While groundbreaking, computer use is still experimental—at times error-prone. We're releasing it early for feedback from developers.
    00:00
    47
  • @AnthropicAI
    Anthropic
    @AnthropicAI
    Oct 22, 2024
    We've built an API that allows Claude to perceive and interact with computer interfaces. This API enables Claude to translate prompts into computer commands. Developers can use it to automate repetitive tasks, conduct testing and QA, and perform open-ended research.
    00:00
    78
  • @AnthropicAI
    Anthropic
    @AnthropicAI
    Oct 22, 2024
    We're trying something fundamentally new. Instead of making specific tools to help Claude complete individual tasks, we're teaching it general computer skills—allowing it to use a wide range of standard tools and software programs designed for people.
    00:00
    43
  • @AnthropicAI
    Anthropic
    @AnthropicAI
    Oct 22, 2024
    Claude 3.5 Sonnet's current ability to use computers is imperfect. Some actions that people perform effortlessly—scrolling, dragging, zooming—currently present challenges. So we encourage exploration with low-risk tasks. We expect this to rapidly improve in the coming months.
    2
  • @AnthropicAI
    Anthropic
    @AnthropicAI
    Even while recording these demos, we encountered some amusing moments. In one, Claude accidentally stopped a long-running screen recording, causing all footage to be lost. Later, Claude took a break from our coding demo and began to peruse photos of Yellowstone National Park.
    00:00
    3:06 PM · Oct 22, 2024·
    530K
    Views
    122
  • @AnthropicAI
    Anthropic
    @AnthropicAI
    Oct 22, 2024
    Beyond computer use, the new Claude 3.5 Sonnet delivers significant gains in coding—an area where it already led the field. Sonnet scores higher on SWE-bench Verified than all available models—including reasoning models like OpenAI o1-preview and specialized agentic systems.
    A comparison table showing benchmark results for Claude 3.5 Sonnet (new) versus Claude 3.5 Sonnet, GPT-4o, and Gemini 1.5 Pro across various tasks like reasoning, coding, and math.
    12
    @AnthropicAI
    Anthropic
    @AnthropicAI
    Oct 22, 2024
    Claude 3.5 Haiku is the next generation of our fastest model. Haiku now outperforms many state-of-the-art models on coding tasks—including the original Claude 3.5 Sonnet and GPT-4o—at the same cost as before. The new Claude 3.5 Haiku will be released later this month.
    A comparison table showing benchmark results for Claude 3.5 Haiku versus Claude 3 Haiku, GPT-4o mini, and Gemini 1.5 Flash across various tasks like reasoning, coding, and math.
    25
    @AnthropicAI
    Anthropic
    @AnthropicAI
    Oct 22, 2024
    We believe these developments will open up new possibilities for how you work with Claude, and we look forward to seeing what you'll create. Read the updates in full:
    Silhouette of person's profile with hand cursor and human head outline
    Introducing computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
    From anthropic.com
    24
  • @jconorgrogan
    Conor
    @jconorgrogan
    Oct 22, 2024
    This is one of the coolest emergent behaviors I've seen
    4