Yandex Cloud
Search
Discuss with expertTry it for free
  • Customer Stories
  • Documentation
  • Blog
  • All Services
    • Cloud Interconnect
    • Cloud Backup
    • Cloud Registry
    • Yandex AI Studio
    • Compute Cloud
    • Object Storage
    • Managed Service for Kubernetes®
    • Yandex BareMetal
    • Smart Web Security
    • Security Deck
    • Managed Service for PostgreSQL
    • Managed Service for ClickHouse®
    • Monium
    • Cloud CDN
    • Network Load Balancer
    • Virtual Private Cloud
    • Cloud DNS
    • Application Load Balancer
    • Yandex Cloud Video
    • Stackland
    • Yandex Cloud Router
    • Yandex Managed Service for Trino
    • Managed Service for MySQL®
    • Managed Service for Valkey™
    • Managed Service for Apache Spark™
    • Yandex StoreDoc
    • Managed Service for OpenSearch
    • Managed Service for Apache Kafka®
    • Data Transfer
    • Yandex MPP Analytics Engine for PostgreSQL
    • Yandex Managed Service for Apache Airflow®
    • Data Processing
    • Yandex MetaData Hub
    • Managed Service for YDB
    • Managed Service for Sharded PostgreSQL
    • Managed Service for YTsaurus
    • Yandex WebSQL
    • DataLens
    • Yandex Search API
    • SpeechSense
    • SpeechKit
    • DataSphere
    • Vision OCR
    • Translate
    • Yandex Identity Hub
    • Key Management Service
    • Certificate Manager
    • Yandex Lockbox
    • Audit Trails
    • SmartCaptcha
    • Cloud Desktop
    • SourceCraft Code Assistant
    • Container Registry
    • Managed Service for GitLab
    • Managed Service for Prometheus®
    • Cloud Functions
    • API Gateway
    • Yandex Cloud Postbox
    • Message Queue
    • Serverless Integrations
    • IoT Core
    • Data Streams
    • Serverless Containers
    • Cloud Notification Service
    • Yandex Query
    • Identity and Access Management
    • Yandex Cloud Console
    • Resource Manager
    • Yandex Cloud Billing
    • Yandex Cloud Quota Manager
    • Cloud Apps
  • System Status
  • Marketplace
    • Featured
    • Infrastructure & Network
    • Data Platform
    • AI for business
    • Security
    • DevOps tools
    • Serverless
    • Monitoring & Resources
  • All Solutions
    • By industry
    • By use case
    • Economics and Pricing
    • Security
    • Technical Support
    • Start testing with double trial credits
    • Cloud credits to scale your IT product
    • Gateway to Russia
    • Cloud for Startups
    • Center for Technologies and Society
    • Yandex Cloud Partner program
    • Price calculator
    • Pricing plans
  • Customer Stories
  • Documentation
  • Blog
© 2026 Direct Cursus Technology L.L.C.
Yandex Data Streams
    • All guides
    • Managing data streams
      • Preparing the environment
      • Creating a data stream
      • Sending data to a stream
      • Reading data from a stream
      • Deleting a stream
  • Access management
  • Pricing policy
  • FAQ
  1. Step-by-step guides
  2. Working with the AWS SDK
  3. Reading data from a stream

Reading data from a stream in the AWS SDK

Written by
Yandex Cloud
Updated at July 15, 2025
View in Markdown
Python

You can get data from a stream using the get_shard_iterator and get_record/get_records methods. When you invoke this method, specify the following parameters:

  • Stream name., e.g., example-stream.
  • ID of the cloud the stream is located in, e.g., b1gi1kuj2dht********.
  • ID of the YDB database containing the stream, e.g., cc8028jgtuab********.

You also need to configure the AWS SDK and assign the service account the yds.viewer role.

To read records from a stream with the above parameters:

  1. Create the stream_get_records.py file and paste the following code to it:

    import boto3
    from pprint import pprint
    import itertools
    
    def get_records(cloud, database, stream_name):
        client = boto3.client('kinesis', endpoint_url="https://yds.serverless.yandexcloud.net")
    
        StreamName = "/ru-central1/{cloud}/{database}/{stream}".format(cloud=cloud,
                                                                     database=database,
                                                                     stream=stream_name)
    
    
        describe_stream_result = client.describe_stream(StreamName=StreamName)
        shard_iterators = {}
    
        shards = [shard["ShardId"] for shard in describe_stream_result['StreamDescription']['Shards']]
    
        for shard_id in itertools.cycle(shards):
            if shard_id not in shard_iterators:
                shard_iterators[shard_id] = client.get_shard_iterator(StreamName=StreamName,
                                                                     ShardId=shard_id,
                                                                     ShardIteratorType='LATEST')['ShardIterator']
               
            record_response = client.get_records(ShardIterator=shard_iterators[shard_id])
            if "Records" in record_response:
                for record in [record for record in record_response["Records"]]:
                    yield record["Data"]
    
            if "NextShardIterator" in record_response:
                shard_iterators[shard_id] = record_response["NextShardIterator"]
    
    
    if __name__ == '__main__':
        for record in get_records(cloud="b1gi1kuj2dht********",
                                  database="cc8028jgtuab********",
                                  stream_name="example-stream"):
            pprint(record)    
            print("The record has been read successfully")
    
  2. Run the program:

    python3 stream_get_records.py
    

    Result:

    The record has been read successfully
    b'{"user_id":"user1","score":100}'
    The record has been read successfully
    b'{"user_id":"user1","score":100}'
    ...
    

Was the article helpful?

Previous
Sending data to a stream
Next
Deleting a stream
© 2026 Direct Cursus Technology L.L.C.