TeleGraphite: Fast and Reliable Telegram Channel Scraper
This repository profile is provided by osrepos.com, an open source repository discovery platform.
Summary
TeleGraphite is a powerful Python tool designed for scraping public Telegram channels efficiently. It allows users to fetch posts, download media, and export all data into structured JSON files. This makes it an ideal solution for data collection, analysis, and archiving Telegram channel content.
Repository Information
Topics
Click on any tag to explore related repositories
Use at your own risk
OSRepos shares public repositories for knowledge and discovery only. Any installation, execution, configuration, or use of code from these repositories is the user's own responsibility. Always review the repository, source code, dependencies, licenses, and security implications before running or installing anything. OSRepos is not responsible for issues, damages, or losses resulting from third-party repositories.
Introduction
TeleGraphite is a fast and reliable Python-based Telegram channel scraper designed to efficiently fetch posts and export them to JSON format. This powerful tool allows users to collect data, download media, and organize content from public Telegram channels with ease. It's an excellent solution for researchers, data analysts, or anyone needing to archive specific channel information.
Key features include:
- Fetching posts from multiple Telegram channels.
- Saving posts as JSON files, including contact exports like emails, phone numbers, and links.
- Downloading and saving media files, such as photos, documents, and videos.
- Deduplicating posts to prevent saving duplicate content.
- Options to run once or continuously with a specified interval.
- Filtering posts by keywords or content type (text-only, media-only).
- Scheduling fetching at specific days and times for automated data collection.
Installation
Getting started with TeleGraphite is straightforward. You can install it either from source or using pip.
From Source
# Clone the repository
git clone https://github.com/hamodywe/telegram-scraper-TeleGraphite.git
cd telegram-scraper-TeleGraphite
# Install the package
pip install -e .
Using pip
pip install telegraphite
Before usage, ensure you have a Telegram API application setup (API ID and API Hash) and a .env file with your credentials. You'll also need a channels.txt file listing the Telegram channels you wish to scrape.
Examples
TeleGraphite offers a flexible command-line interface for various scraping scenarios.
Basic Usage
Fetch posts once and exit:
telegraphite once
Fetch posts continuously with a 1-hour interval:
telegraphite continuous --interval 3600
Advanced Options
Fetch 20 posts from each channel and save to a custom directory:
telegraphite once --limit 20 --data-dir custom_data
Fetch only posts containing specific keywords:
telegraphite once --keywords announcement important news
Run continuously on specific days and times:
telegraphite continuous --days monday wednesday friday --times 09:00 18:00
You can also use a YAML configuration file for more complex setups, allowing you to define filters, schedules, and other options.
Why use TeleGraphite?
TeleGraphite stands out as a robust solution for Telegram data extraction due to its comprehensive feature set and ease of use. It provides unparalleled flexibility for collecting, organizing, and analyzing public Telegram channel content. Whether you need to archive historical posts, monitor ongoing discussions, or extract specific data points like contact information, TeleGraphite simplifies the process. Its ability to download media, deduplicate content, and run on a schedule makes it a powerful tool for automated data acquisition, saving significant time and effort.
Links
- GitHub Repository: hamodywe/telegram-scraper-TeleGraphite
- Telegram API Application: my.telegram.org
Related repositories
Similar repositories that may be relevant next.

Benchmark Radar: A Living Database for AI Benchmarks and Evaluation
September 29, 2026
Benchmark Radar is an extensive open-source project that tracks over 20,710 AI benchmark, evaluation, dataset, and data-quality records from 37 public sources. It provides daily updates, linked evidence, and tools for researchers and developers to discover and analyze AI benchmarks. This project is essential for anyone needing to stay current with AI evaluation trends and model performance.

Pydantic AI Harness: Enhancing Your AI Agents with Robust Capabilities
September 28, 2026
Pydantic AI Harness is the official capability and harness library for Pydantic AI, designed to extend agents for complex, long-running tasks. It provides a modular system of "capabilities" for functionalities like file system interaction, web research, memory, and sub-agent delegation. This library enables developers to build sophisticated and durable AI agents with ease.

Bernstein: Open-Source Governance and Orchestration for AI Agents
September 28, 2026
Bernstein is an open-source framework designed for the governance and orchestration of AI agents, allowing users to define rules declaratively. It enforces these policies and generates verifiable, replayable records of all agent activities. This Python-based solution provides a robust layer for managing complex AI agent workflows with transparency and accountability.

Meshtastic-MCP: AI Tooling for Meshtastic Device Control and Testing
September 27, 2026
Meshtastic-MCP provides an MCP server and agent skills designed for AI tooling to discover, drive, observe, and test Meshtastic devices and applications. It offers a comprehensive suite of capabilities, from portable device control to advanced hardware-free end-to-end testing and replay functionalities. This project aims to streamline the development and testing of Meshtastic ecosystems.
Source repository
Open the original repository on GitHub.
9 counted GitHub visits