RDF System Management Manual

ManualsBrandsHP ManualsServerHP NonStop G-Series

361

362

363

364

365

366

367

368

369

370

Table Of Contents

RDF System Management Manual
What’s New in This Manual
About This Manual
1 Introducing RDF
- RDF Subsystem Overview
  - Unplanned Outages With ZLT
  - Unplanned Outages Without ZLT
  - Tips for Executing Fast Business Takeover Operations
  - Planned Outages
  - Features
  - User Interfaces
  - Tasks
- RDF Processes
  - Primary System Processes
  - Backup System Processes
- RDF Operations
  - Monitor Process
  - Extractor Process
  - Receiver Process
  - RDFNET Process
  - Updater Processes
  - Purger Process
- Reciprocal and Chain Replication
- Available Types of Replication to Multiple Backup Systems
  - RDF Control Subvolume
- Triple Contingency
- Loopback Configuration (Single System)
- Online Product Initialization
- Online Database Synchronization
- Online Dumps
- Subvolume- and File-Level Replication
- Shared Access DDL Operations
- EMS Support
- SMF Support
- RTD Warning Thresholds
- Process-Lockstep Operation
- Support for Network Transactions
- RDF and NonStop SQL/MX
- Zero Lost Transactions (ZLT)
- Monitoring RDF Entities With ASAP
2 Preparing the RDF Environment
- Configuring Hardware for RDF Operations
- Preparing Software and Database Files for RDFOperations
- Using SMF With RDF
  - Configuring an SMF Environment on the Primary System
  - Configuring an SMF Environment on the Backup RDF System
3 Installing and Configuring RDF
- Preparing the Primary System
  - Stopping the Software
  - Preparing the Tables and Files
- Preparing the Backup System
  - Synchronizing the Primary and Backup Databases
  - Backing Up Application Programs and Files
- Installing RDF
- Initializing and Configuring TMF
  - TMF Subsystem Not Running Previously
  - TMF Subsystem Running Previously
- Initializing and Configuring RDF
- Enabling RDF Operations
4 Operating and Monitoring RDF
- Running RDFCOM
- Running RDFSCAN
- Performing Routine Operational Tasks
5 Managing RDF
- Recovering From File System Errors
- Handling Disk Space Problems
- Responding to Operational Failures
- Stopping RDF
- Restarting RDF
- Carrying Out a Planned Switchover
- Takeover Operations
- Reading the Backup Database
- Access to Backup Databases in a Consistent State
- RDF and NonStop SQL/MP DDL Operations
  - Performing Non-Shared Access DDL Operations
  - Performing Shared Access DDL Operations
- RDF and NonStop SQL/MX Operations
- Backing Up Image Trail Files
- Making Online Dumps With Updaters Running
- Doing FUP RELOAD Operations With Updaters Running
- Exception File Optimization
- Switching Disks on Updater UPDATEVOLUMES
6 Maintaining the Databases
- Understanding Database States
- Backing Up Altered Database Structures
  - NonStop SQL/MP or NonStop SQL/MX Databases
  - Enscribe Databases
- Resynchronizing Databases
7 Online Database Synchronization
- Overview
- Synchronizing Entire Databases Online
  - Considerations When Synchronizing Entire Databases
  - Example of Synchronizing An Entire Database Online
- Synchronizing Selected Database Portions Online
  - Overview
  - Partial Database Synchronization Issues
- Phases of Online Database Synchronization
  - Extractor Phases
  - Updater Phase 2
- Extractor Restart Considerations During Online Database Synchronization
- When Is Online Database Synchronization Complete?
  - Extractor Messages
  - Updater Messages
8 Entering RDFCOM Commands
- Command Description Elements
- File Names and Process Identifiers
- Command Overview
9 Entering RDFSCAN Commands
- About the EMS Log
- Command Description Elements
- AT
- DISPLAY
- EXIT
- FILE
- HELP
- LIST
- LOG
- MATCH
- NOLOG
- SCAN
10 Triple Contingency
- What Is It?
- What’s Required?
- How Does It Work?
- Hardware Requirements
- Software Requirements
- The RETAINCOUNT Configuration Parameter
- The COPYAUDIT Command
- COPYAUDIT Restartability
- Using ZLT to Achieve Triple Contingency Protection for Auxiliary Audit Trails
  - Triple Contingency Without ZLT
  - Using ZLT to Achieve the Same Protection
- Summary
11 Subvolume- and File-Level Replication
- INCLUDE Clauses
- EXCLUDE Clauses
- Wildcard Character (*)
  - Within Subvolume Names
  - Within Filenames
- INCLUDE and EXCLUDE Processing
- Error Checking
- Performance Ramifications
- Summary Examples
12 Auxiliary Audit Trails
- Auxiliary Extractor
- Auxiliary Receiver
- Configuring Extractors and Receivers
  - Error conditions
- Configuring Image Trails
- Configuring Updaters
  - Error Conditions
- STOP TMF Ramifications
- Takeover Ramifications
- Usage of Master and Auxiliary Audit Trails
- Using Expand Multi-CPU Paths
13 Network Transactions
- Configuration Changes
- RDF Network Control Files
- Normal RDF Processing Within a Network Environment
- RDF Takeovers Within a Network Environment
- Network Configurations and Shared Access NonStop SQL/MP DDL Operations
- Network Validation and Considerations
- RDF Re-Initialization in a Network Environment
  - Network Master Subsystem Initialization
  - Non Network Master Subsystem Initialization
- RDF Networks and ABORT or STOP RDF Operations
- RDF Networks and Stop-Update-to-Time Operations
- Sample Configurations
  - Sample Network Master Configuration
  - Sample Non Network Master Configuration
- RDFCOM STATUS Display
14 Process-Lockstep Operation
- Starting a Lockstep Operation
- The DoLockstep Procedure
- The Lockstep Transaction
- RDF Lockstep File
- Multiple Concurrent Lockstep Operations
- The Lockstep Gateway Process
- Disabling Lockstep
- Reenabling Lockstep
- Lockstep Performance Ramifications
- Lockstep and Auxiliary Audit Trails
- Lockstep and Network Transactions
- Lockstep Operation Event Messages
15 NonStop SQL/MX and RDF
- Including and Excluding SQL/MX Objects
- Obtaining ANSI Object Names From Updater Event Messages
- Creating NonStop SQL/MX Primary and Backup Databases from Scratch
- Creating a NonStop SQL/MX Backup Database From an Existing Primary Database
- Online Database Synchronization With NonStop SQL/MX Objects
  - Creating the Fuzzy Copy on the Primary System
  - Creating the Fuzzy Copy on the Backup System
- Offline Synchronization for a Single Partition
- Online Synchronization for a Single Partition
- Correcting Incorrect NonStop SQL/MX Name Mapping
- Consideration for Creating Backup Tables
- Restoring to a Specific Location
  - Example
- Comparing NonStop SQL/MX Tables
16 Zero Lost Transactions (ZLT)
- How It Works
- Using CommitHoldMode
- Hardware Setup
- Assigning CPUs on the Standby System
- RDF Configuration Attributes
- ZLT Takeover Operations
- Recovering the Primary System After an RDF ZLT Takeover
- ZLT and RDF Networks
- STOP TMF Operations
  - During Normal Operations
  - During ZLT Takeover Processing
- SQL Shared Access DDL Operations
A RDF Command Summary
- RDFSCAN Commands
- File Names and Process Identifiers
B Additional Reference Information
- Default Configuration Parameters
- Sample Configuration File
- RDFSNOOP Utility
- RDF System Files
- RDF File Codes
C Messages
- About the Message Descriptions
- RDF Messages
- RDFCOM Messages
- RDFSCAN Messages
D Operational Limits
E Using ASAP
- Architectural Overview
- Installation
- Auto Discovery
- Monitoring Specific RDF Environments
- Adding and Removing RDF Environments
- Version Compatibility
- RDF Metrics Reported by ASAP
Index

Network Transactions

HP NonStop RDF System Management Manual—524388-003

13-8

Communication Failures During Phase 3 Takeover

Processing

have to go through 60 minutes of data to determine what must be undone due to data

missing on the system that had fallen behind.

A variation of the first example is that no extractors have fallen behind, but you have 25

systems in your RDF network. In such a case, phase 3 processing may take many

additional seconds because data must be checked for so many different systems in

order to determine what network data might be missing from the various systems in the

RDF network.

Communication Failures During Phase 3 Takeover Processing

If one RDF subsystem is unable to reach the backup system of another RDF subsytem

during phase 3 processing, phase 3 processing stalls until the communication line

comes back up. This can lengthen the overall duration of takeover operations on all

backup systems. Should this type of stall occur, the RDF subsystem issues an event

message alerting operators to the situation.

Takeover Delays and Purger Restarts

During phase 3 purger work, the network master needs information from the other

purger processes in the RDF network, and, during the latter part of phase 3

processing, the non-network master purgers need information from the purger of the

network master. When a purger process is waiting for information from another purger,

it waits for up to 60 seconds, during which time it does not respond to certain requests

(such as STATUS RDF). After a purger has waited 60 seconds, it quits the operation

and restarts. This allows the purger to read the $RECEIVE file, respond to messages

that have been waiting for replies, and then retry phase 3 processing.

Takeover Restartability

As has always been the case, the RDFCOM TAKEOVER command is restartable.

Therefore, if a takeover operation terminates prematurely for any reason on any

system in an RDF network, it can be restarted.

Takeover and File Recovery

When a takeover operation completes in an RDF network environment, the purger logs

two events: one reports a safe MAT position (indicating that all committed data up to

that location was successfully applied to the backup database), and the second (888 or

858) reports whether or not a File Recovery position is available for use on the primary

system. The RDF event 888 reports that a File Recovery is available and it includes

the exact sno and rba to be used for a File Recovery operation on the primary system.

If, however, “kept-commits” have been encountered during phase 2 processing, a File

Recovery position is not available; this is reported in RDF event 858. Note that this

last situation will never occur in an RDF/ZLT environment because, with RDF/ZLT, a

File Recovery position is always available.