Attachment

53 KB Posted

Attached to
Mega 4 Automated Litigation Support Federal contract opportunity
Solicitation number
DJJI-12-RFP-0783
Issued by
Department of Justice Offices Boards and Divisions Justice Management Division

About this file

Attachment (7) Accuracy Standards

Text of this file

Attachment No. 5, Accuracy Standards

January, 2001 Page 1

ACCURACY STANDARDS

I. Accuracy: Definition And Formula

Subject to the conditions stated below, the Contractor shall ensure that the document coding and data reduction tasks are performed at a level of quality such that:

l. all recorded information is accurate;

2. all relevant information is recorded; and,

3. no irrelevant information is recorded.

"Accuracy" shall be determined on the basis of a character count. The formula for determining the accuracy of each batch of document coding or its associated data tape shall be:

A - (B + C + D) A X 100 where: A = all characters which should be recorded B = all recorded characters which are inaccurate

C = all relevant characters which are not recorded D = all irrelevant characters which are recorded.

As the Statement of Work provides, the COTR shall determine for each Task Order the size of the coding/keying "batch. "

For purposes of example, suppose the coded data associated with the documents from CD XY01 is designated as a coding batch, and suppose that in sampling this batch, 35 documents are selected for examination.

Upon review, 12,351 characters should have been entered for these 35 documents (not counting the characters for Document Number and Document Date). If the actual coding for these documents had 23 relevant characters which were inaccurately coded ("B") in the formula above), 18 relevant characters which were omitted ("C") and 65 irrelevant characters recorded ("D"), the accuracy of the coding batch would be:

12,351 - (23 + 18 + 65) 12,351 X 100 = 99.14%

This coding batch would be rejected on the basis of the accuracy requirements and definition of "characters" given in Sections II and III below.

II. Document Coding: Definition of "Characters"

For the purposes of determination of coding accuracy, characters shall be counted as illustrated below:

EXAMPLE 1: (Document Date field)

*DT 760101

6 = 6 characters

EXAMPLE 2: (Document Title, illustrating a variable length text field)

*TI REPORT ON EXPENDITURES

20 = 20 characters

EXAMPLE 3: (Document Type field, illustrating "check-offs" for valid field types)

January, 2001 Page 2

*TY LET ; MEM ; REP ; CHR

1 + 1 = 2 characters

Only the coded characters, as defined and illustrated above, are counted.

For litigation document databases, the Document Number and Document Date fields must be 100% accurate. A single error in either of these fields shall constitute grounds for rejection of the entire batch. The minimum acceptance standard for the remaining fields will be 99.95% based on a character count, as defined and illustrated above.

For all databases other than litigation document database, the COTR will define the field which constitutes the "unique identifying record number" and any other primary sort fields (for instance "Deposition Date" in the Case Management System). For each of these databases, the "unique identifying record number" and the other primary sort fields (no more than two other sort fields) must be 100% accurate. A single error in such a field constitutes grounds for rejection of the entire batch. The minimum acceptance standard for the remaining fields will be 99. 95% based on a count of characters.

All determinations of accuracy will be made by comparing the document coding forms with the associated source documents for a statistically valid random sampling of documents from the batch. Coding deliverables will not be considered accepted until 30 days after the successful loading of all DCFs in a given batch, including correction of errors detected during the loading process.

III. Keying/Key Verification: Definition of "Characters"

For the purposes of billing and determination of data reduction accuracy, characters shall be counted as illustrated below:

EXAMPLE l: (Document Date Field)

*DT 760101

l + l (tab)+ 6 + l (enter) = 9 characters

EXAMPLE 2: (Author field, illustrating all subfields completed)

*AU JONES ,LP % ABC MINING

l + l (tab) + 8 + 1 + 10 + l (enter) = 22 characters

EXAMPLE 3: (Title field)

*TI REPORT ON EXPENDITURES

l + l (tab) + 22 + l (enter) = 24 characters

EXAMPLE 4: (Document Type field, illustrating "checkoffs" for valid field values)

*TY LET ;MEM ;REP ;CHR

l + l (tab) + 3 + l (delim. ) + 3 + l (enter) = 10 characters

Only the keyed characters, as defined and illustrated above, are counted; key verification characters are not counted.

For litigation document databases, the Document Number and Document Date fields must be 100% accurate.

A single error in either of these fields shall constitute grounds for rejection of the entire batch. The minimum

January, 2001 Page 3 acceptance standard for the remaining fields will be 99.98% based on a character count, as defined and illustrated above.

For all databases other than litigation document database, the COTR will define the field which constitutes the "unique identifying record number" and any other primary sort fields (for instance "Deposition Date" in the Case Management System). For each of these databases, the "unique identifying record number" and the other primary sort fields (no more than two other sort fields) must be 100% accurate. A single error in such a field constitutes grounds for rejection of the entire batch. The minimum acceptance standard for the remaining fields will be 99.98% based on a count of characters.

All determinations of accuracy will be made by comparing the keyed data with the associated document coding forms for a statistically valid random sampling of document coding forms from the batch. Keying deliverables will not be considered accepted until 30 days after the successful loading of all DCFs in a given batch, including correction of all errors detected during the loading process.

IV. Overall Database Accuracy Standards

For purposes of determining overall database accuracy with respect to the documents, characters shall be counted as illustrated below:

EXAMPLE 1: (Document Type Field)

MEM 3 Characters CHR 3 Characters 6 Characters

EXAMPLE 2: (Author Field)

ANDREWS,J SAS 12 Characters

Only characters of data present in the database are counted; tags, "repeat" numbers, trailing blanks, etc. are not counted.

For litigation document databases, the Document Number and Document Date fields must be 100% accurate.

A single error in either of these fields shall constitute grounds for rejection of the entire batch. The minimum acceptance standard for the remaining fields, when compared with the information in the documents themselves, shall be 99.90% based on a character count, as illustrated above.

For databases other than litigation document databases, the COTR will define the fields which constitute the "unique identifying record number" and the other primary sort fields (no more than two other sort fields). For each of these data bases, the "unique identifying record number" and other primary sort fields must be 100% accurate. The minimum acceptance standard for the remaining fields will be 99.90% based on a character count, as illustrated above.

All determinations of accuracy will be made by comparing printouts of the database records with a statistically valid sampling of the hard copies of the documents which they represent. Coding, keying, tape loading, and error correction deliverables will not be considered accepted until 30 days after the successful loading of all DCFs in a given batch, including correction of errors detected during the loading process.

V. Accuracy of OCRed Text

For purposes of calculating the accuracy of OCRed text, characters and characters in error will be counted as per the following examples:

Example 1:

[Original Transcript Text]

January, 2001 Page 4

12 more if they look like it would be complicated. In the

Total character count = 55.

[OCRed Text For Same Line]

12 more hey look like it would be cohplicated. In. the

Total missing characters = 3

Total incorrect characters = 1

Total unnecessary/irrelevant characters = 1.

Accuracy will be measured by visually comparing OCR output against the original page. Accuracy for this line would be calculated as follows:

55 - (3 + 1 + 1) x 100 = 90.9% .

Notes:

Extra spaces between words in the OCRed text are not counted as errors. However, an extra space in the middle of a word is counted as an error.

Special characters inserted by the OCRing device to flag potential OCRing errors are counted as erroneous characters.

Page 1
Page 2
Page 3
Page 4

Other files for this federal contract opportunity

Other files attached to Mega 4 Automated Litigation Support, newest first.
File Type Posted
Industry_Q As.pdf PDF
amendment_0002.pdf PDF
Attachment_(9)_Labor_Category_Descriptions_(Amendment_0002).pdf PDF
Atachment_(1)_Pricing_Tables_(Amendment_0002).xlsx XLSX spreadsheet
Attachment_(2)_Adjustment_Factors_(Amendment_0002).xlsx XLSX spreadsheet
Attachment_(8)_Anticipated_Workload_at_Time_of_Award_(Amendment_0002).pdf PDF
RFP_(Amendment_0002).pdf PDF
Atachment —
Amendment 0001.pdf PDF
Attachment —
Attachment —
Attachment —
Attachment —
SF-33Form.pdf PDF
Attachment —
Attachment —
Attachment —
RFP.docx DOCX document
Attachment —
Attachment —
Attachment —
Show all 21

On GovTribe

Work with this file on GovTribe

  • Download the original file
  • Contacts named in this file
  • Similar government files
  • Ask GovTribe AI about this file

File details come from the government source that posted it. Updated .