本文へ移動
cccskills
無料GitHub で公開

threads-user-posts

Fetches public posts from a Threads user's profile page, extracting post text, engagement metrics, and media info from SSR-embedded JSON. Use when user asks to scrape Threads posts, get someone's Threads feed, pull posts from a Threads account, collect Threads content by username, download Threads user posts, extract Threads profile posts, monitor a Threads user's activity, retrieve recent posts from a Threads handle, gather Threads post data for a creator, or fetch all posts from a specific Threads profile.

インストール方法を見る

含まれるファイル(3)

  • SKILL.md8.9 KB
  • scripts/extract-posts.py3.1 KB
  • scripts/scroll-load-more.py793 B

SKILL.md(原文)

インストールする前に、エージェントに与えられる指示の中身を確認できます。

Threads — User Posts

username → list of public posts with engagement metrics (SSR + scroll pagination)

Language

All process output to user (progress updates, process notifications) follows the user's language.

Objective

Extract public posts from a given Threads user's profile page using SSR-embedded JSON data, with optional scroll-triggered pagination to load more posts.

Prerequisites

  • A browser is open and connected via browser-act
  • Login recommended: Threads enforces a login wall on profile pages in most regions as of mid-2026. A logged-in browser session is required for reliable access. Unauthenticated access via US-based proxies may work in some configurations but is not guaranteed.

Pre-execution Checks

1. Tool Readiness

If browser-act has been confirmed available in the current session → skip this step.

Invoke browser-act via Skill tool to load usage. If installation or configuration issues arise, follow its guidance to resolve then retry.

2. Login Verification

If login status for Threads has been confirmed in the current session → skip this step.

Otherwise: navigate https://www.threads.com and check page state:

  • User avatar or username visible at top → logged in, continue
  • "Log in" / "Sign up" buttons visible → not logged in, inform the user that login is needed, assist completing the login flow via the Threads login page

If not logged in and user cannot log in → the capability may still work on some proxy/IP configurations; attempt navigate https://www.threads.com/@{username} and check if profile content loads. If the page redirects to /login/, terminate and inform the user that a logged-in session is required.

Capability Components

This Skill's operational boundary = what the user can manually do in their browser. It only reads data already displayed to the user on the page. JS code is encapsulated in Python files under the scripts/ directory, invoked via eval "$(python scripts/xxx.py {params})". $(...) is bash syntax; it is recommended to use the bash tool for execution.

SSR: Extract posts from profile page (initial load)

Navigate to the user's profile page first, then extract all posts embedded in the initial page HTML:

  1. navigate https://www.threads.com/@{username}
  2. wait stable
  3. eval "$(python scripts/extract-posts.py)"

Output example:

{
  "posts": [
    {
      "id": "3936653356768062022",          // post internal ID (pk)
      "code": "DKmzN8wJzFG",               // post short code for URL
      "url": "https://www.threads.com/@zuck/post/DKmzN8wJzFG",  // direct post link
      "text": "Excited to share...",        // post text content, null if no caption
      "taken_at": 1780953646,              // Unix timestamp of post creation
      "like_count": 3583,                  // number of likes
      "reply_count": 3039,                 // number of direct replies
      "repost_count": 232,                 // number of reposts
      "quote_count": 114,                  // number of quote posts
      "reshare_count": 460,                // number of reshares
      "is_reply": false,                   // true if this post is a reply to another
      "media_type": 19,                    // 1=photo, 2=video, 8=carousel, 19=text-only
      "has_media": false,                  // true if post contains image/video/carousel
      "user": {
        "pk": "63055343223",              // Threads-specific user ID
        "username": "zuck",
        "full_name": "Mark Zuckerberg",
        "is_verified": true
      }
    }
  ],
  "count": 4,                             // number of posts in this batch
  "page_info": {
    "end_cursor": "QVFES...",             // cursor for next page (null when no more)
    "has_next_page": true,                // whether more posts are available
    "has_previous_page": false,
    "start_cursor": null
  }
}

Error handling: If error: true is returned, check that the profile page was fully loaded (navigate and wait stable completed without timeout). If mediaData not found, the page may have redirected to a login wall or the username is invalid.

DOM: Scroll to trigger more posts

When page_info.has_next_page is true, scroll the page to trigger GraphQL auto-load:

eval "$(python scripts/scroll-load-more.py)"

Then wait stable and read the new batch from network traffic:

network requests --type xhr,fetch --filter threads.com

Find the POST https://www.threads.com/graphql/query request(s) that appeared after the scroll. Read the response:

network request <id>

The response body contains data.mediaData.edges[] with the same post structure as the SSR extraction output. Each edge has node.thread_items[0].post with the same field layout (id/pk, code, caption.text, taken_at, like_count, text_post_app_info.direct_reply_count, etc.).

Error handling: If no new graphql/query requests appear after scrolling, the user may have reached the login wall (~15 posts without authentication) or all posts have loaded. Check for a "Log in to see more" message on screen via screenshot.

Composite: Full profile post collection (SSR + scroll pagination)

To collect all available posts for a user:

  1. navigate https://www.threads.com/@{username} → wait stable
  2. eval "$(python scripts/extract-posts.py)" → collect initial posts, note page_info
  3. While page_info.has_next_page == true: a. eval "$(python scripts/scroll-load-more.py)" b. wait stable c. network requests --type xhr,fetch --filter threads.com → locate new graphql/query POST d. network request <id> → extract data.mediaData.edges[] posts e. Check data.mediaData.page_info.has_next_page in response to decide whether to continue
  4. Stop when has_next_page == false or login wall encountered (~15 posts without login)

Pagination

DOM Pagination: Scroll #scrollview via eval "$(python scripts/scroll-load-more.py)", then wait stable and read new POST /graphql/query response from network traffic. Termination: page_info.has_next_page == false in the GraphQL response, or login wall visible on screen (approximately 15 posts without authentication).

Success Criteria

result.count >= 1 and result.posts[0].id != null and result.posts[0].user.username != null

Known Limitations

  • Login wall: As of mid-2026, Threads redirects unauthenticated profile page requests to the login page in most regions. A logged-in browser session is required for reliable use; unauthenticated access works only on specific proxy/IP configurations.
  • Authenticated access: up to approximately 15 posts per profile before hitting a soft pagination limit; further posts require additional scroll cycles.
  • Private accounts: profile page shows no posts even for logged-in users who don't follow the account.
  • Deleted or suspended accounts return mediaData not found error.
  • taken_at is a Unix timestamp; convert to readable date with new Date(taken_at * 1000).toISOString().

Execution Efficiency

  • Batch orchestration: Write a bash script to loop through the command templates serially within a single session; do not parallelize within one browser. Add 2-3 second intervals between profile navigations to avoid rate limiting. For higher throughput, distribute usernames across multiple parallel browser sessions.
  • Test before batch execution: After writing a batch script, test with 1-2 usernames first to verify the script runs correctly; only then run the full batch.
  • Reduce redundant pre-operations: When collecting multiple profiles in the same session, the session state is preserved across navigations — no need to reinitialize.
  • Error resumption: Save results per username during batch processing; on failure, resume from the breakpoint rather than starting over.

Experience Notes

Path: {working-directory}/browser-act-skill-forge-memories/threads-scraper-threads-user-posts.memory.md

Before execution: If the file exists, read it first — it records unexpected situations encountered during past executions (e.g., a strategy has become ineffective); adjust strategy order accordingly.

After execution: If an unexpected situation is encountered (strategy became ineffective, page redesigned, anti-scraping upgraded, better path discovered), append a line: {YYYY-MM-DD}: {what happened} → {conclusion}

Normal execution does not write to the file. Do not record what keywords were used or how many results were returned — those are task outputs, not experience.

レビュー

まだレビューはありません。使ってみた感想をお寄せください。

同じリポジトリのスキル

概要と使いどころ

Extracts comprehensive wholesale product data from 1688.com product detail pages: title, tiered pricing, SKU variants with dimensions/weight, product images, seller info, shop scores, buyer protection, cross-border flags, product attributes, coupon/promotion data, and review stats. Use when user mentions 1688, 1688.com, wholesale China, alibaba wholesale, B2B China sourcing, Chinese wholesale scraper, 1688 product scrape, 1688 offer, 1688 detail, extract 1688 data, pull 1688 listings, get wholesale price, 1688 supplier info, factory stats 1688, 1688 SKU variants, 1688 product attributes, 1688 shop score, DSR score 1688, 1688 buyer protection, 1688 cross-border, 1688 dropship. Also applies to: scraping bulk product data from 1688 by offer ID list, monitoring 1688 supplier metrics, extracting 1688 pricing tiers for resale analysis.

日本語の概要は準備中です。原文の説明を表示しています。

browser-act/skills6,1292026年8月24日 更新

Fetches complete Airbnb listing details for a given numeric listing ID via the internal GraphQL API, returning title, room type, description, amenities, photos, coordinates, city, house rules, highlights, ratings, review count, bedroom configuration, and property overview. Use when user mentions Airbnb listing details, Airbnb property info, Airbnb room details, get Airbnb listing data, Airbnb amenities list, Airbnb house rules, Airbnb property description, Airbnb detail page scraper, Airbnb rooms detail, Airbnb property page data, Airbnb listing info, fetch Airbnb room details, pull Airbnb listing.

日本語の概要は準備中です。原文の説明を表示しています。

browser-act/skills6,1292026年8月24日 更新

Extracts Airbnb accommodation search results from a destination query via SSR-embedded data, returning listing ID, URL, name, coordinates, rating, price, photos, and badge info for each result, plus pagination cursors for multi-page retrieval. Use when user mentions Airbnb search results, Airbnb listings, vacation rental search, short-term rental listings, scrape Airbnb, get Airbnb data, find rentals on Airbnb, Airbnb destination search, Airbnb property list, Airbnb stays search, Airbnb accommodation results, pull Airbnb listings, collect Airbnb search data, Airbnb scraper, Airbnb search page extraction, Airbnb search by destination.

日本語の概要は準備中です。原文の説明を表示しています。

browser-act/skills6,1292026年8月24日 更新

Amazon Alexa for Shopping Q&A automation: submits questions to Amazon's Alexa/Rufus AI shopping assistant and collects response text; supports optional keyword search context (navigate to search results page before asking for category-specific answers). Use when user mentions Amazon Alexa, Rufus, Amazon shopping assistant, Amazon AI chat, ask Amazon, Amazon Q&A, automate Alexa questions, Rufus chatbot, Amazon assistant automation, collect Alexa responses, bulk question submission to Amazon, keyword search context, category research. Also applies to extracting Amazon product recommendations from conversational AI, automating repeated queries to Amazon's AI shopping feature, collecting Alexa shopping responses at scale, or market research within a specific product category.

日本語の概要は準備中です。原文の説明を表示しています。

browser-act/skills6,1292026年8月24日 更新

This skill helps users extract structured product details from Amazon using a specific ASIN (Amazon Standard Identification Number). Use this skill when the user asks to get Amazon product details by ASIN, lookup Amazon product title and price using ASIN, extract Amazon product ratings and reviews count for a specific ASIN, check Amazon product availability and current price, get Amazon product description and features via ASIN, enrich product catalog with Amazon data using ASIN, monitor Amazon product price changes for specific ASINs, retrieve Amazon product brand and material information, fetch Amazon product images and specifications by ASIN, validate Amazon ASIN and get product metadata.

日本語の概要は準備中です。原文の説明を表示しています。

browser-act/skills6,1292026年8月24日 更新

This skill helps users extract structured best-selling product data from Amazon via the BrowserAct API. Agent should proactively apply this skill when users express needs like search for best selling products on Amazon, extract Amazon product data based on keywords, find top rated Amazon products, monitor Amazon competitor prices and sales, discover trending products on Amazon marketplace, extract Amazon product titles prices and ratings, gather Amazon product sales volume for market research, search Amazon best sellers in specific region, collect Amazon product reviews and promotion details, analyze Amazon product availability and badges, get Amazon product data for market analysis.

日本語の概要は準備中です。原文の説明を表示しています。

browser-act/skills6,1292026年8月24日 更新

browser-act のスキルをすべて見る

このスキルの問題を報告する