MySQL - History Tables 이력관리 / 히스토리 테이블

MySQL - History Tables

Keeping a history of your data can be immensely useful, such as for reverting silly mistakes, or for auditing purposes. This tutorial will show you a really simple way to achieve this in a generic manner that can be applied to any table. You will be able to see what data and when, as well as return to any specific revision or point in time quickly and easily.

Preparation

I am assuming you already have a dev database to work with.

Run the following statements to create a table of data that we are going to demonstrate with throghout this tutorial.

CREATE TABLE `user_comments` (
    `id` int UNSIGNED NOT NULL AUTO_INCREMENT,
    `comment` text NOT NULL,
    `author_id` int NOT NULL,
    `modified_timestamp` timestamp DEFAULT CURRENT_TIMESTAMP ON UPDATE CURRENT_TIMESTAMP,
    PRIMARY KEY (`id`)
) ENGINE=InnoDB DEFAULT CHARSET=utf8;

INSERT INTO `user_comments`
(`comment`, `author_id`) VALUES
("hello world", 1),
("foo bar", 2);
Copy to clipboard

If your table's don't already have the timestamp, or author fields, then I recommend that you add them. Without the timestamp field, you will only be able to go back to a specific revision number rather than a certain point in time. Without the author_id field, you will not know who made the changes.

Steps

The first thing we need to do is clone the original table's schema to create our history table.

CREATE TABLE `user_comments_history` LIKE `user_comments`;

By using a suffix of _history rather than a prefix, you keep your history tables beside the ones they track in your database list.

From this point on, there are two main ways of structuring your history tables. For both of these options, a new field will be added to the history table which will be called history_id.

In option 1, the new history_id field will be the primary key, and id in the history table will just be a data field containing the value of and referencing the ID in the main table. This has the advantage of being very simple for copying data rows to/from the history table as the data in the columns remain exactly the same.

In option 2, the id field in the history table will remain as the primary key, and a new field (history_id) will reference the id in the original table. The advantage of this is that the schema doesn't change as id remains the primary key, but you now have to move data between the id and history_id fields when moving rows. You may find this option simpler if you rename history_id to row_id or primary_table_id. I am just keeping the name the same between the two options for this tutorial.

Option 1

Run the following steps to alter the history table to how we need it.

ALTER TABLE `user_comments_history`
MODIFY COLUMN `id` INT UNSIGNED NOT NULL;

ALTER TABLE `user_comments_history` DROP PRIMARY KEY;

ALTER TABLE `user_comments_history`
ADD COLUMN `history_id` INT UNSIGNED NOT NULL;

ALTER TABLE `user_comments_history`
ADD CONSTRAINT PRIMARY KEY (`history_id`);

ALTER TABLE `user_comments_history`
MODIFY `history_id` INT UNSIGNED NOT NULL AUTO_INCREMENT;

It seems long-winded but unfortunately adding the primary key to a new column has to be done in that many steps.

Adding A Foreign Key

It may be a good idea to add a foreign key to enforce the relationship between the history table and the rows in the original table.

ALTER TABLE `user_comments_history`
ADD CONSTRAINT fk_id FOREIGN KEY (id) REFERENCES `user_comments`(id) ON UPDATE CASCADE ON DELETE CASCADE;

Updating A Row

Now when we wish to update a row, we need to insert into the history table first. The code below shows how to do this in a single transaction so that both have to go through or neither. We are going to change the first row's comment from hello world to hello earth.

경축! 아무것도 안하여 에스천사게임즈가 새로운 모습으로 재오픈 하였습니다.
어린이용이며, 설치가 필요없는 브라우저 게임입니다.
https://s1004games.com

START TRANSACTION;

# Insert the current row into the history table
INSERT INTO `user_comments_history` (`id`, `comment`, `author_id`, `modified_timestamp`)
    SELECT
      `id` as `id`,
      `comment` as `comment`,
      `author_id` as `author_id`,
      `modified_timestamp` as `modified_timestamp`
    FROM `user_comments`
    WHERE `comment` = 'hello world';

# Now run the update
UPDATE `user_comments`
SET `comment`='hello earth'
WHERE `comment` = 'hello world';

# Commit the transaction
COMMIT;

Restoring To A Point In Time

If you know that your data was fine on the 11th of July 2016 and you want to retrieve the data from that point in time, then just use the following query (assuming we want the history of row 1 in the primary table):

SELECT * FROM `user_comments_history`
WHERE `id`= 1
AND `modified_timestamp` < "2016-11-30"
LIMIT 1
ORDER BY `modified_timestamp` DESC

However, to change back to that point in time, then run the following transaction:

START TRANSACTION;

# Insert the current row into the history table before reverting
INSERT INTO `user_comments_history` (`id`, `comment`, `author_id`, `modified_timestamp`)
    SELECT
      `id` as `id`,
      `comment` as `comment`,
      `author_id` as `author_id`,
      `modified_timestamp` as `modified_timestamp`
    FROM `user_comments`
    WHERE `id`= 1

# Now run the update
UPDATE `user_comments` dest,
(
  SELECT * FROM `user_comments_history`
  WHERE `id`= 1
  AND `modified_timestamp` < "2016-11-30"
  ORDER BY `modified_timestamp` DESC
  LIMIT 1
) src
SET
dest.comment = src.comment,
dest.author_id = dest.author_id
WHERE dest.`id`= 1;

# Commit the transaction
COMMIT;

Option 2

If you believe that the id of every table needs to be the primary key, you can do the following instead:

ALTER TABLE `user_comments_history`
ADD COLUMN `history_id` INT UNSIGNED NOT NULL;

It may be a good idea to add a foreign key to enforce the relationship between the history table and the rows in the original table.

ALTER TABLE `user_comments_history`
ADD CONSTRAINT fk_id FOREIGN KEY (history_id) REFERENCES `user_comments`(id);

Updating A Row

Now when we wish to update a row, we need to insert into the history table first. The code below shows how to do this in a single transaction so that both have to go through or neither. We are going to change the first row's comment from hello world to hello earth.

START TRANSACTION;

# Insert the current row into the history table
INSERT INTO `user_comments_history` (`history_id`, `comment`, `author_id`, `modified_timestamp`)
    SELECT `id` as `history_id`, `comment` as `comment`, `author_id` as `author_id`, `modified_timestamp` as `modified_timestamp`
    FROM `user_comments`
    WHERE `comment` = 'hello world'
;

# Update the row int he primary table
UPDATE `user_comments`
SET `comment`='hello earth'
WHERE `comment` = 'hello world';

# Commit the transaction
COMMIT;

Restoring To A Point In Time

If you know that your data was fine on the 11th of July 2016 and you want to retrieve the data from that point in time, then just use the following query (assuming we want the history of row 1 in the primary table):

SELECT * FROM `user_comments_history`
WHERE `history_id`= 1
AND `modified_timestamp` < "2016-11-30"
LIMIT 1
ORDER BY `modified_timestamp` DESC

However, to change back to that point in time, then run the following transaction:

START TRANSACTION;

# Insert the current row into the history table before reverting
INSERT INTO `user_comments_history` (`history_id`, `comment`, `author_id`, `modified_timestamp`)
    SELECT
      `id` as `history_id`,
      `comment` as `comment`,
      `author_id` as `author_id`,
      `modified_timestamp` as `modified_timestamp`
    FROM `user_comments`
    WHERE `id`= 1

# Now run the update
UPDATE `user_comments` dest,
(
  SELECT * FROM `user_comments_history`
  WHERE `history_id`= 1
  AND `modified_timestamp` < "2016-11-30"
  ORDER BY `modified_timestamp` DESC
  LIMIT 1
) src
SET
dest.comment = src.comment,
dest.author_id = dest.author_id
WHERE dest.`id`= 1;

# Commit the transaction
COMMIT;

References

Appendix

To Use A Foreign Key or Not

It may be a good idea to add a foreign key to enforce the relationship between the history table and the rows in the original table. However if you do this, then it must be the case that if you DELETE a row from the primary table, its entire history is also removed. If you need to keep the history such a situation, then do not implement the foreign key, but your application layer will need to ensure to cascade any updates that occur to the primary table's ID. Unlike the rest of the columns, you cannot let the IDs diverge because otherwise you do not know which row in the primary table that the history table relates to. Based on my experience, there is usually no reason for the ID of a row to change but its something to be aware of.

Depending on your circumstances, sometimes it is easier for rows to have a "state" field that can be altered to mark the row as "deleted" than to actually delete the row. For example if a user deletes their account, you may wish to put into a "deleted" state rather than actually removing their data. That way the data is there if the user changes their mind at a later date. Such a scenario would allow you to keep the foreign key.

 

[출처] https://blog.programster.org/mysql-history-tables

 

 

본 웹사이트는 광고를 포함하고 있습니다.
광고 클릭에서 발생하는 수익금은 모두 웹사이트 서버의 유지 및 관리, 그리고 기술 콘텐츠 향상을 위해 쓰여집니다.
번호 제목 글쓴이 날짜 조회 수
공지 오라클 기본 샘플 데이터베이스 졸리운_곰 2014.01.02 86968
공지 [SQL컨셉] 서적 "SQL컨셉"의 샘플 데이타 베이스 SAMPLE DATABASE of ORACLE 가을의 곰을... 2013.02.10 79230
공지 [G_SQL] Sample Database 가을의 곰을... 2012.05.20 95974
78 [Kafka] Kafka 한번 살펴보자... Quickstart file 졸리운_곰 2021.06.18 1315
77 Java Kafka Producer, Consumer 예제 구현 Java를 이용하여 Kafka Producer와 Kakfa Consumer를 구현해보자. file 졸리운_곰 2021.06.18 1172
76 Beginner’s Guide to Understand Kafka file 졸리운_곰 2021.06.18 1575
75 [Kafka] Kafka 설치/실행 및 테스트 file 졸리운_곰 2021.06.18 1067
74 [java] [kafka] [Kafka] 개념 및 기본예제 file 졸리운_곰 2021.06.16 2171
73 Getting started with Apache Kafka in Python file 졸리운_곰 2020.09.10 2262
72 [Kafka] 다운로드 및 Quick Start file 졸리운_곰 2020.09.07 1900
71 [Kafka] 기본 개념잡기 file 졸리운_곰 2020.09.07 1720
70 Flume Integration with Kafka file 졸리운_곰 2019.04.16 2087
69 빅데이터: 플럼(Flume) 토폴로지 설계 file 졸리운_곰 2019.04.16 1604
68 실시간 처리를 위한 분산 메시징 시스템 카프카(Kafka) file 졸리운_곰 2018.05.12 1413
67 Flume과 Kafka를 사용한 초당 100만개 로그 수집 테스트 file 졸리운_곰 2018.05.12 1402
66 웹 크롤링 / web crwaling / web scraping / 웹 스크래핑 file 졸리운_곰 2017.07.09 1849
65 빅데이터 단지 몇퍼센트의 예측 정확성을 위하여 장애로 가득찬 빅데이터 시스템을 도입하여야 하는가에 대한 의문! file 졸리운_곰 2017.03.20 1612
64 빅데이터: 플럼(Flume) 토폴로지 설계 file 졸리운_곰 2017.03.20 1422
63 [실시간 분석 시스템] Apache Flume를 활용한 데이터 수집(1) file 졸리운_곰 2017.03.06 1324
62 [실시간 분석 시스템] 데이터 수집 #2 Apache Sqoop을 활용하여 RDBMS 데이터 수집(2) file 졸리운_곰 2017.03.06 1388
61 [실시간 분석 시스템] 데이터 수집 #2 Apache Sqoop을 활용하여 RDBMS 데이터 수집(1) file 졸리운_곰 2017.03.06 1109
60 [실시간 분석 시스템] 데이터 수집 #1 오픈 소스 수집기 비교 file 졸리운_곰 2017.03.06 1752
59 [실시간 분석 시스템] 일단 데이터 들여다 보기 file 졸리운_곰 2017.03.06 1948
대표 김성준 주소 : 경기 용인 분당수지 U타워 등록번호 : 142-07-27414
통신판매업 신고 : 제2012-용인수지-0185호 출판업 신고 : 수지구청 제 123호 개인정보보호최고책임자 : 김성준 sjkim70@stechstar.com
대표전화 : 010-4589-2193 [fax] 02-6280-1294 COPYRIGHT(C) stechstar.com ALL RIGHTS RESERVED