RDF System Management Manual

ManualsBrandsHP ManualsServerHP NonStop G-Series

Table Of Contents

RDF System Management Manual
What’s New in This Manual
About This Manual
1 Introducing RDF
- RDF Subsystem Overview
  - Unplanned Outages With ZLT
  - Unplanned Outages Without ZLT
  - Tips for Executing Fast Business Takeover Operations
  - Planned Outages
  - Features
  - User Interfaces
  - Tasks
- RDF Processes
  - Primary System Processes
  - Backup System Processes
- RDF Operations
  - Monitor Process
  - Extractor Process
  - Receiver Process
  - RDFNET Process
  - Updater Processes
  - Purger Process
- Reciprocal and Chain Replication
- Available Types of Replication to Multiple Backup Systems
  - RDF Control Subvolume
- Triple Contingency
- Loopback Configuration (Single System)
- Online Product Initialization
- Online Database Synchronization
- Online Dumps
- Subvolume- and File-Level Replication
- Shared Access DDL Operations
- EMS Support
- SMF Support
- RTD Warning Thresholds
- Process-Lockstep Operation
- Support for Network Transactions
- RDF and NonStop SQL/MX
- Zero Lost Transactions (ZLT)
- Monitoring RDF Entities With ASAP
2 Preparing the RDF Environment
- Configuring Hardware for RDF Operations
- Preparing Software and Database Files for RDFOperations
- Using SMF With RDF
  - Configuring an SMF Environment on the Primary System
  - Configuring an SMF Environment on the Backup RDF System
3 Installing and Configuring RDF
- Preparing the Primary System
  - Stopping the Software
  - Preparing the Tables and Files
- Preparing the Backup System
  - Synchronizing the Primary and Backup Databases
  - Backing Up Application Programs and Files
- Installing RDF
- Initializing and Configuring TMF
  - TMF Subsystem Not Running Previously
  - TMF Subsystem Running Previously
- Initializing and Configuring RDF
- Enabling RDF Operations
4 Operating and Monitoring RDF
- Running RDFCOM
- Running RDFSCAN
- Performing Routine Operational Tasks
5 Managing RDF
- Recovering From File System Errors
- Handling Disk Space Problems
- Responding to Operational Failures
- Stopping RDF
- Restarting RDF
- Carrying Out a Planned Switchover
- Takeover Operations
- Reading the Backup Database
- Access to Backup Databases in a Consistent State
- RDF and NonStop SQL/MP DDL Operations
  - Performing Non-Shared Access DDL Operations
  - Performing Shared Access DDL Operations
- RDF and NonStop SQL/MX Operations
- Backing Up Image Trail Files
- Making Online Dumps With Updaters Running
- Doing FUP RELOAD Operations With Updaters Running
- Exception File Optimization
- Switching Disks on Updater UPDATEVOLUMES
6 Maintaining the Databases
- Understanding Database States
- Backing Up Altered Database Structures
  - NonStop SQL/MP or NonStop SQL/MX Databases
  - Enscribe Databases
- Resynchronizing Databases
7 Online Database Synchronization
- Overview
- Synchronizing Entire Databases Online
  - Considerations When Synchronizing Entire Databases
  - Example of Synchronizing An Entire Database Online
- Synchronizing Selected Database Portions Online
  - Overview
  - Partial Database Synchronization Issues
- Phases of Online Database Synchronization
  - Extractor Phases
  - Updater Phase 2
- Extractor Restart Considerations During Online Database Synchronization
- When Is Online Database Synchronization Complete?
  - Extractor Messages
  - Updater Messages
8 Entering RDFCOM Commands
- Command Description Elements
- File Names and Process Identifiers
- Command Overview
9 Entering RDFSCAN Commands
- About the EMS Log
- Command Description Elements
- AT
- DISPLAY
- EXIT
- FILE
- HELP
- LIST
- LOG
- MATCH
- NOLOG
- SCAN
10 Triple Contingency
- What Is It?
- What’s Required?
- How Does It Work?
- Hardware Requirements
- Software Requirements
- The RETAINCOUNT Configuration Parameter
- The COPYAUDIT Command
- COPYAUDIT Restartability
- Using ZLT to Achieve Triple Contingency Protection for Auxiliary Audit Trails
  - Triple Contingency Without ZLT
  - Using ZLT to Achieve the Same Protection
- Summary
11 Subvolume- and File-Level Replication
- INCLUDE Clauses
- EXCLUDE Clauses
- Wildcard Character (*)
  - Within Subvolume Names
  - Within Filenames
- INCLUDE and EXCLUDE Processing
- Error Checking
- Performance Ramifications
- Summary Examples
12 Auxiliary Audit Trails
- Auxiliary Extractor
- Auxiliary Receiver
- Configuring Extractors and Receivers
  - Error conditions
- Configuring Image Trails
- Configuring Updaters
  - Error Conditions
- STOP TMF Ramifications
- Takeover Ramifications
- Usage of Master and Auxiliary Audit Trails
- Using Expand Multi-CPU Paths
13 Network Transactions
- Configuration Changes
- RDF Network Control Files
- Normal RDF Processing Within a Network Environment
- RDF Takeovers Within a Network Environment
- Network Configurations and Shared Access NonStop SQL/MP DDL Operations
- Network Validation and Considerations
- RDF Re-Initialization in a Network Environment
  - Network Master Subsystem Initialization
  - Non Network Master Subsystem Initialization
- RDF Networks and ABORT or STOP RDF Operations
- RDF Networks and Stop-Update-to-Time Operations
- Sample Configurations
  - Sample Network Master Configuration
  - Sample Non Network Master Configuration
- RDFCOM STATUS Display
14 Process-Lockstep Operation
- Starting a Lockstep Operation
- The DoLockstep Procedure
- The Lockstep Transaction
- RDF Lockstep File
- Multiple Concurrent Lockstep Operations
- The Lockstep Gateway Process
- Disabling Lockstep
- Reenabling Lockstep
- Lockstep Performance Ramifications
- Lockstep and Auxiliary Audit Trails
- Lockstep and Network Transactions
- Lockstep Operation Event Messages
15 NonStop SQL/MX and RDF
- Including and Excluding SQL/MX Objects
- Obtaining ANSI Object Names From Updater Event Messages
- Creating NonStop SQL/MX Primary and Backup Databases from Scratch
- Creating a NonStop SQL/MX Backup Database From an Existing Primary Database
- Online Database Synchronization With NonStop SQL/MX Objects
  - Creating the Fuzzy Copy on the Primary System
  - Creating the Fuzzy Copy on the Backup System
- Offline Synchronization for a Single Partition
- Online Synchronization for a Single Partition
- Correcting Incorrect NonStop SQL/MX Name Mapping
- Consideration for Creating Backup Tables
- Restoring to a Specific Location
  - Example
- Comparing NonStop SQL/MX Tables
16 Zero Lost Transactions (ZLT)
- How It Works
- Using CommitHoldMode
- Hardware Setup
- Assigning CPUs on the Standby System
- RDF Configuration Attributes
- ZLT Takeover Operations
- Recovering the Primary System After an RDF ZLT Takeover
- ZLT and RDF Networks
- STOP TMF Operations
  - During Normal Operations
  - During ZLT Takeover Processing
- SQL Shared Access DDL Operations
A RDF Command Summary
- RDFSCAN Commands
- File Names and Process Identifiers
B Additional Reference Information
- Default Configuration Parameters
- Sample Configuration File
- RDFSNOOP Utility
- RDF System Files
- RDF File Codes
C Messages
- About the Message Descriptions
- RDF Messages
- RDFCOM Messages
- RDFSCAN Messages
D Operational Limits
E Using ASAP
- Architectural Overview
- Installation
- Auto Discovery
- Monitoring Specific RDF Environments
- Adding and Removing RDF Environments
- Version Compatibility
- RDF Metrics Reported by ASAP
Index

Introducing RDF

HP NonStop RDF System Management Manual—524388-003

1-5

Unplanned Outages Without ZLT

Without ZLT functionality, it is possible for some committed transactions to be lost

during an unplanned outage. When the RDF TAKEOVER command is issued, any

transaction whose final outcome is unknown on the backup system is backed out of the

backup database. One or more transactions might have committed on the primary

system, but, before the extractor could read and send the associated audit data to the

backup system, the primary system failed. Loss of audit data in this manner typically

involves no more than a fraction of a second.

If the primary system is unexpectedly brought down because of a disaster, the

outcome of some transactions might never be known, as illustrated in Table 1-1.

In the example illustrated in Table 1-1, a disaster has brought down the primary system

immediately after the commit record for transaction 100 was written to the MAT, but

before the RDF extractor process was able to send the commit record to the backup

system. For transaction 101, a single update was logged in the MAT and sent to the

backup system, but the primary system was brought down before the transaction was

completed.

When the command for a takeover is issued, the updater processes treat all

transactions whose outcomes are not known as aborted transactions. In this scenario,

only the changes related to transactions known with certainty to have been committed

on the primary system are left in the backup database. Therefore, in the example

illustrated in Table 1-1, the audit information associated with transactions 100 and 101

is backed out of the backup database.

Typically, the extractor process sends audit information to the backup system within a

second after it has been written to the MAT on the primary system, so a minimum

number of transactions are lost when a disaster brings down the primary system.

Table 1-1. Audit Information At the Time of a Primary System Failure

Primary database updates

(Sequence in master audit trail file)

Updates sent to the backup

(Sequence in image trail file)

TRANS100—Update 1 TRANS100—Update 1

TRANS100—Update 2 TRANS100—Update 2

TRANS100—Update 10 TRANS100—Update 10

TRANS101—Update 1 TRANS101—Update 1

TRANS100—Commit record

(Primary system fails)