RDF System Management Manual

ManualsBrandsHP ManualsServerHP NonStop G-Series

161

162

163

164

165

166

167

168

169

170

Table Of Contents

RDF System Management Manual
What’s New in This Manual
About This Manual
1 Introducing RDF
- RDF Subsystem Overview
  - Unplanned Outages With ZLT
  - Unplanned Outages Without ZLT
  - Tips for Executing Fast Business Takeover Operations
  - Planned Outages
  - Features
  - User Interfaces
  - Tasks
- RDF Processes
  - Primary System Processes
  - Backup System Processes
- RDF Operations
  - Monitor Process
  - Extractor Process
  - Receiver Process
  - RDFNET Process
  - Updater Processes
  - Purger Process
- Reciprocal and Chain Replication
- Available Types of Replication to Multiple Backup Systems
  - RDF Control Subvolume
- Triple Contingency
- Loopback Configuration (Single System)
- Online Product Initialization
- Online Database Synchronization
- Online Dumps
- Subvolume- and File-Level Replication
- Shared Access DDL Operations
- EMS Support
- SMF Support
- RTD Warning Thresholds
- Process-Lockstep Operation
- Support for Network Transactions
- RDF and NonStop SQL/MX
- Zero Lost Transactions (ZLT)
- Monitoring RDF Entities With ASAP
2 Preparing the RDF Environment
- Configuring Hardware for RDF Operations
- Preparing Software and Database Files for RDFOperations
- Using SMF With RDF
  - Configuring an SMF Environment on the Primary System
  - Configuring an SMF Environment on the Backup RDF System
3 Installing and Configuring RDF
- Preparing the Primary System
  - Stopping the Software
  - Preparing the Tables and Files
- Preparing the Backup System
  - Synchronizing the Primary and Backup Databases
  - Backing Up Application Programs and Files
- Installing RDF
- Initializing and Configuring TMF
  - TMF Subsystem Not Running Previously
  - TMF Subsystem Running Previously
- Initializing and Configuring RDF
- Enabling RDF Operations
4 Operating and Monitoring RDF
- Running RDFCOM
- Running RDFSCAN
- Performing Routine Operational Tasks
5 Managing RDF
- Recovering From File System Errors
- Handling Disk Space Problems
- Responding to Operational Failures
- Stopping RDF
- Restarting RDF
- Carrying Out a Planned Switchover
- Takeover Operations
- Reading the Backup Database
- Access to Backup Databases in a Consistent State
- RDF and NonStop SQL/MP DDL Operations
  - Performing Non-Shared Access DDL Operations
  - Performing Shared Access DDL Operations
- RDF and NonStop SQL/MX Operations
- Backing Up Image Trail Files
- Making Online Dumps With Updaters Running
- Doing FUP RELOAD Operations With Updaters Running
- Exception File Optimization
- Switching Disks on Updater UPDATEVOLUMES
6 Maintaining the Databases
- Understanding Database States
- Backing Up Altered Database Structures
  - NonStop SQL/MP or NonStop SQL/MX Databases
  - Enscribe Databases
- Resynchronizing Databases
7 Online Database Synchronization
- Overview
- Synchronizing Entire Databases Online
  - Considerations When Synchronizing Entire Databases
  - Example of Synchronizing An Entire Database Online
- Synchronizing Selected Database Portions Online
  - Overview
  - Partial Database Synchronization Issues
- Phases of Online Database Synchronization
  - Extractor Phases
  - Updater Phase 2
- Extractor Restart Considerations During Online Database Synchronization
- When Is Online Database Synchronization Complete?
  - Extractor Messages
  - Updater Messages
8 Entering RDFCOM Commands
- Command Description Elements
- File Names and Process Identifiers
- Command Overview
9 Entering RDFSCAN Commands
- About the EMS Log
- Command Description Elements
- AT
- DISPLAY
- EXIT
- FILE
- HELP
- LIST
- LOG
- MATCH
- NOLOG
- SCAN
10 Triple Contingency
- What Is It?
- What’s Required?
- How Does It Work?
- Hardware Requirements
- Software Requirements
- The RETAINCOUNT Configuration Parameter
- The COPYAUDIT Command
- COPYAUDIT Restartability
- Using ZLT to Achieve Triple Contingency Protection for Auxiliary Audit Trails
  - Triple Contingency Without ZLT
  - Using ZLT to Achieve the Same Protection
- Summary
11 Subvolume- and File-Level Replication
- INCLUDE Clauses
- EXCLUDE Clauses
- Wildcard Character (*)
  - Within Subvolume Names
  - Within Filenames
- INCLUDE and EXCLUDE Processing
- Error Checking
- Performance Ramifications
- Summary Examples
12 Auxiliary Audit Trails
- Auxiliary Extractor
- Auxiliary Receiver
- Configuring Extractors and Receivers
  - Error conditions
- Configuring Image Trails
- Configuring Updaters
  - Error Conditions
- STOP TMF Ramifications
- Takeover Ramifications
- Usage of Master and Auxiliary Audit Trails
- Using Expand Multi-CPU Paths
13 Network Transactions
- Configuration Changes
- RDF Network Control Files
- Normal RDF Processing Within a Network Environment
- RDF Takeovers Within a Network Environment
- Network Configurations and Shared Access NonStop SQL/MP DDL Operations
- Network Validation and Considerations
- RDF Re-Initialization in a Network Environment
  - Network Master Subsystem Initialization
  - Non Network Master Subsystem Initialization
- RDF Networks and ABORT or STOP RDF Operations
- RDF Networks and Stop-Update-to-Time Operations
- Sample Configurations
  - Sample Network Master Configuration
  - Sample Non Network Master Configuration
- RDFCOM STATUS Display
14 Process-Lockstep Operation
- Starting a Lockstep Operation
- The DoLockstep Procedure
- The Lockstep Transaction
- RDF Lockstep File
- Multiple Concurrent Lockstep Operations
- The Lockstep Gateway Process
- Disabling Lockstep
- Reenabling Lockstep
- Lockstep Performance Ramifications
- Lockstep and Auxiliary Audit Trails
- Lockstep and Network Transactions
- Lockstep Operation Event Messages
15 NonStop SQL/MX and RDF
- Including and Excluding SQL/MX Objects
- Obtaining ANSI Object Names From Updater Event Messages
- Creating NonStop SQL/MX Primary and Backup Databases from Scratch
- Creating a NonStop SQL/MX Backup Database From an Existing Primary Database
- Online Database Synchronization With NonStop SQL/MX Objects
  - Creating the Fuzzy Copy on the Primary System
  - Creating the Fuzzy Copy on the Backup System
- Offline Synchronization for a Single Partition
- Online Synchronization for a Single Partition
- Correcting Incorrect NonStop SQL/MX Name Mapping
- Consideration for Creating Backup Tables
- Restoring to a Specific Location
  - Example
- Comparing NonStop SQL/MX Tables
16 Zero Lost Transactions (ZLT)
- How It Works
- Using CommitHoldMode
- Hardware Setup
- Assigning CPUs on the Standby System
- RDF Configuration Attributes
- ZLT Takeover Operations
- Recovering the Primary System After an RDF ZLT Takeover
- ZLT and RDF Networks
- STOP TMF Operations
  - During Normal Operations
  - During ZLT Takeover Processing
- SQL Shared Access DDL Operations
A RDF Command Summary
- RDFSCAN Commands
- File Names and Process Identifiers
B Additional Reference Information
- Default Configuration Parameters
- Sample Configuration File
- RDFSNOOP Utility
- RDF System Files
- RDF File Codes
C Messages
- About the Message Descriptions
- RDF Messages
- RDFCOM Messages
- RDFSCAN Messages
D Operational Limits
E Using ASAP
- Architectural Overview
- Installation
- Auto Discovery
- Monitoring Specific RDF Environments
- Adding and Removing RDF Environments
- Version Compatibility
- RDF Metrics Reported by ASAP
Index

Managing RDF

HP NonStop RDF System Management Manual—524388-003

5-24

Takeover Failure

For RDF network takeover considerations, see Section 13, Network Transactions.

For super fast takeover, see Tips for Executing Fast Business Takeover Operations in

section 1.

Takeover Failure

If a double CPU failure occurs and the receiver process pair or an updater process pair

fails during a takeover operation, you can resume the operation just by entering the

TAKEOVER command through RDFCOM again. You can ascertain that a takeover

operation failed by issuing a STATUS RDF command and getting a response such as

the following:

STATUS RDF (\RDF04 -> \RDF05) is NOT running

A partial RDF TAKEOVER has completed

Also, a takeover failure generates a 725 event in the EMS log.

Monitor Considerations

Whether the RDF monitor was started when the initial TAKEOVER command was

executed or not, this process is always started when the TAKEOVER command is

reissued.

Updater Considerations

When the purger shuts down at the end of the takeover operation, it examines the

context record of each updater process to determine if that updater has processed all

applicable audit data through the end-of-file in the image trail. If all updaters have

processed through the end-of-file, the purger logs a 724 message to the EMS event

log, indicating that the takeover operation completed successfully. But if it determines

that one or more updaters have terminated prematurely, the purger logs RDF Message

726 for each updater that failed and then logs RDF Message 725, a general message

indicating that the takeover operation did not complete successfully. If these messages

appear in the EMS event log, you must reissue the TAKEOVER command.

Takeover and File Recovery

After you initiate a takeover, it is possible that the last committed transactions did not

make it to the backup system (meaning that the backup and primary databases are not

synchronized).

If the takeover completes on the backup system, the purger logs an RDF event 888

specifying a MAT position (sno, rba). Subsequently, when the primary system is once

again online and you are ready to switch the applications back to the primary, you first

initiate a TMF file recovery command with the TOMATPOSITION option on the primary

system specifying the logged MAT position from the 888 event. TMF restores the