WEB-BASED CONTENT
SCENARIO: CALIF. RECALL ELECTION
The California Recall Election Project of the California Digital Library captured a collection of web sites from the historic 2003 California gubernatorial recall election.
· Assurance/permissions obtained after content had been collected
· No links outside of domain were followed
· Depth: ??
· Content from:
official candidate sites
official California State sites
California county and local sites
political interest groups
commentary/humor sites specific to recall
media sites, including news sites, captured
· Types of notification and response:
descriptive web page outlining scope and purpose
identification of owners, mailing of ‘opt-out’ notification
procedures established for take-down
notification of site removal
· Crawl composed of both single and periodic captures
· No assurances for site content
· Images, logos, etc., may not be owned by sited
· Consent is “passive”
· Inline content, if outside of site boundaries, may not be captured
· Fair use, in terms of copying large parts of web sites, has not been established
· Records should be maintained for date or dates site is crawled
· Store permission and agreements, preferably with resource
· Store information about dates of capture and attempts to capture
· Record technology used and settings regarding how crawl was executed for compliance under robot exclusion protocol
· Log take-down requests and actions with proff of action
· Maintain records showing due diligence in seeking approval