Investigating Deleted or Archived Tweets with Legendary OSINT: A Complete Guide

Legendary OSINT includes multiple third-party tools for recovering deleted or archived tweets, including Wayback Tweets, BirdHunt, Nitter, snscrape, and Twint, all curated in the repository's social media documentation.

Legendary OSINT serves as a comprehensive knowledge base for open-source intelligence analysts, cataloging reliable utilities for social media investigations. When investigating deleted or archived tweets, analysts can reference the People Search & Social Media documentation to discover specialized tools that resurrect removed content or query historical indexes. These utilities are not built into the repository itself but are rigorously vetted and organized for immediate operational use.

Curated Tools for Deleted Tweet Recovery

The repository maintains a curated list of utilities specifically for X (formerly Twitter) investigations in docs/people-social.md. These tools address different aspects of tweet recovery, from archive retrieval to historical indexing.

Wayback Tweets

Wayback Tweets is the primary tool listed for retrieving deleted content from the Internet Archive. According to the source code in docs/people-social.md (lines 135-136), this web application pulls tweet content directly from archived snapshots, allowing analysts to view text, timestamps, and metadata for posts no longer available on the live platform. The service provides a REST API endpoint at https://waybacktweets.streamlit.app/api/tweet/{tweet_id} for programmatic access.

BirdHunt

BirdHunt appears in the same documentation section (lines 132-133) as a historical tweet search service. Unlike real-time scrapers, BirdHunt exposes past tweet data that may have been removed from X's public interface but remains indexed in specialized databases. This tool is particularly valuable for investigating accounts that have purged their historical content.

Nitter

Nitter is documented (line 136) as an alternative front-end for X that enables scraping without triggering API rate limits. While primarily an interface replacement, Nitter is OSINT-friendly and pairs effectively with automated scraping scripts to capture tweet pages before they are deleted, creating private archives for later analysis.

snscrape and Twint

For programmatic collection, docs/people-social.md (lines 129-131) lists snscrape and Twint as Python-based scrapers that operate without API keys. These tools enable analysts to pull tweets, profiles, and hashtags directly from X's web interface, creating local datasets that preserve content even if the original posts are subsequently deleted.

Practical Implementation Examples

The following Python implementations demonstrate how to integrate these tools into an investigative workflow, combining direct scraping with historical archive retrieval.

Scraping Live Tweets with snscrape

Use snscrape to collect recent tweet data before it potentially disappears. This approach captures JSON-formatted tweet metadata via command-line execution:


# Requires: pip install snscrape

import subprocess
import json

def fetch_recent_tweets(username, limit=20):
    # snscrape outputs JSON lines; capture via subprocess

    cmd = [
        "snscrape", "--json", "--max-results", str(limit),
        f"twitter-user:{username}"
    ]
    result = subprocess.run(cmd, capture_output=True, text=True)
    tweets = [json.loads(line) for line in result.stdout.splitlines()]
    return tweets

tweets = fetch_recent_tweets("example_user")
print(f"Fetched {len(tweets)} tweets")
print(tweets[0]["content"])

Querying the Wayback Tweets API

For content already deleted from X, query the Wayback Tweets API to retrieve archived versions. This method accesses the Internet Archive's snapshots through a streamlined interface:

import requests

def get_wayback_tweet(tweet_id):
    url = f"https://waybacktweets.streamlit.app/api/tweet/{tweet_id}"
    resp = requests.get(url, timeout=10)
    if resp.status_code == 200:
        return resp.json()  # Contains text, created_at, user, etc.

    else:
        raise ValueError(f"Tweet {tweet_id} not found in archive")

# Replace with the numeric ID of the target tweet

archived = get_wayback_tweet("1234567890123456789")
print("Archived tweet text:", archived["text"])

These implementations illustrate two distinct investigative approaches: proactive collection using snscrape to preserve current data, and reactive recovery via Wayback Tweets for content that has already been removed.

Summary

  • Legendary OSINT catalogs five primary tools for investigating deleted or archived tweets in docs/people-social.md: Wayback Tweets, BirdHunt, Nitter, snscrape, and Twint.
  • Wayback Tweets and BirdHunt specialize in historical recovery, while snscrape and Twint enable real-time data preservation.
  • Nitter provides an alternative front-end for bypassing API restrictions during scraping operations.
  • The repository functions as a knowledge base rather than a toolkit, directing analysts to external, verified utilities.
  • Python integrations using subprocess for snscrape and requests for Wayback Tweets API enable automated investigation pipelines.

Frequently Asked Questions

Does Legendary OSINT include built-in tools for recovering deleted tweets?

No, Legendary OSINT does not contain built-in recovery tools. Instead, the repository serves as a curated knowledge base that documents and links to reliable third-party services like Wayback Tweets and BirdHunt, as organized in docs/people-social.md.

How does Wayback Tweets retrieve content that has been deleted from X?

Wayback Tweets queries the Internet Archive's Wayback Machine for cached snapshots of specific tweet URLs. When a tweet is archived before deletion, this service extracts the text, metadata, and media references from the historical snapshot, making it accessible even after the original post is removed from X's servers.

What is the difference between snscrape and Twint for tweet investigation?

Both are Python-based scrapers listed in docs/people-social.md (lines 129-131) that operate without API keys, but they use different technical approaches. snscrape is actively maintained and scrapes X's web interface for current data, while Twint (now largely unmaintained) was historically used for similar purposes. Analysts currently favor snscrape for reliability in production environments.

Where can I find the complete list of social media investigation tools in the repository?

The complete curated list resides in docs/people-social.md, specifically within the "X (former Twitter)" section spanning lines 129-136. This file contains direct links to Wayback Tweets, BirdHunt, Nitter, and the Python scrapers, along with contextual descriptions of their investigative applications.

Have a question about this repo?

These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:

Share the following with your agent to get started:
curl -s "https://instagit.com/install.md"

Works with
Claude Codex Cursor VS Code OpenClaw Any MCP Client

Maintain an open-source project? Get it listed too →