Kunkel on Artificial Intelligence, Automation, and Proletarianization of the Legal Profession

Rebecca Kunkel (Rutgers Law) has posted “Artificial Intelligence, Automation, and Proletarianization of the Legal Profession” (Creighton Law Review, Vol. 56, 2022) on SSRN. Here is the abstract:

Recent advances in computer programming, broadly categorized as “artificial intelligence,” (“Al”) have renewed debates over machines as viable replacements for human lawyers. Some prominent lawyers and legal scholars now adhere to a vision of the future heavily seasoned with Silicon Valley-style techno-utopianism: the legal profession may endure but only in a form in which it would be almost unrecognizable today, while legal innovators will need to immerse themselves in the possibilities opened up by artificial intelligence in order to survive. For others, the view of artificial intelligence and its potential application to law is more limited, as they argue for the impossibility of automating many essential aspects of legal service. These views share key assumptions about the nature of Al technology: that technological development follows its own course and that the widespread adoption of technologies is primarily determined by objective measures of efficacy. This essay offers an alternate Marxian account of legal Al which places it in the larger history of automation and proletarianization.

Chalkidis on ChatGPT Cannot (Yet) Pass LexGLUE Benchmark

Ilias Chalkidis (University of Copenhagen) has posted “ChatGPT May Pass the Bar Exam Soon, but Has a Long Way to Go for the LexGLUE Benchmark” on SSRN. Here is the abstract:

Following the hype around OpenAI’s ChatGPT conversational agent, the last straw in the recent development of Large Language Models (LLMs) that demonstrate emergent unprecedented zero-shot capabilities, we audit the latest OpenAI’s GPT-3.5 model, ‘gpt-3.5-turbo’, the first available ChatGPT model, in the LexGLUE benchmark in a zero-shot fashion providing examples in a templated instruction-following format. The results indicate that ChatGPT achieves an average micro-F1 score of 49.0% across LexGLUE tasks, surpassing the baseline guessing rates. Notably, the model performs exceptionally well in some datasets, achieving micro-F1 scores of 62.8% and 70.1% in the ECtHR B and LEDGAR datasets, respectively. The code base and model predictions are available at https://github.com/coastalcph/zeroshot_lexglue.

Asay on the DMCA’s Anti-Circumvention Provisions

Clark D. Asay (Brigham Young Law) has posted “An Empirical Study of the DMCA’s Anti-Circumvention Provisions” on SSRN. Here is the abstract:

The DMCA has been a flashpoint during most of its twenty-five-year existence. One of the most controversial parts of the DMCA is Section 1201. Among other things, Section 1201 prohibits third parties from circumventing certain controls to copyrighted content or trafficking in tools that enable circumvention of technological controls. However, despite its nearly quarter-of-a-century lifespan, we know very little about Section 1201 empirically. While certain aspects of the broader DMCA have received empirical assessments, Section 1201 has not. Our understanding of Section 1201 is largely based on anecdotal evidence, in the form of leading opinions from historically prominent copyright circuits. But this anecdotal evidence is hardly a solid basis for ongoing discussions about how Section 1201 is performing and whether it needs revising. In this Article, we seek to address these and other issues.

To do so, we conducted a broad-based search of Westlaw to collect every issued opinion, whether reported or not, where a court purported to apply some part of Section 1201. We then reviewed these cases to glean as much useful information about Section 1201 as possible. This review led to a number of important and, in some cases, surprising results. First, Section 1201 opinions are a relative rarity. In the nearly quarter of a century since the DMCA’s enactment, we could find only a little over 200 opinions, with only about sixty of those being published. The average number of opinions during the DMCA’s existence has been around nine annually, which pales in comparison to other types of copyright cases. Second, despite the Second Circuit receiving much attention in anecdotal accountings of Section 1201, courts within it issue Section 1201 opinions infrequently. The Ninth Circuit is the dominant Section 1201 court, both in terms of citations to its opinions and overall number of opinions, and the Sixth and Eleventh Circuits both issue more Section 1201 opinions than the Second Circuit. This result stands in contrast to other types of copyright litigation, where the Second Circuit is a behemoth. Third, the most common subject matter in dispute in Section 1201 cases is computer software, followed distantly by audiovisual material such as movies. Music stands in last place, showing up in only a couple issued opinions. Debates at the time of the DMCA’s enactment were informed by widespread fears of copyright infringement relating to digital music and other types of digital content. Yet Section 1201 has resulted in but few litigations involving those subject matters. Fourth, suits and defaults against individuals happen frequently in the Section 1201 context, with courts often assessing large statutory damages against those individuals. As we discuss in the paper, this result raises important equity issues. Fifth, despite Section 1201 including a number of statutory exceptions, these exceptions basically never make their way into issued opinions. Fair use, too, only infrequently enters courts’ Section 1201 discussions. This means, effectively, that the primary way to escape Section 1201 liability is through administrative exceptions granted by the Library of Congress on a triennial basis. But as we shall see, this process has significant holes. Finally, plaintiffs disproportionately win Section 1201 cases. This result is somewhat bloated because of the frequency of defaults against individuals. Setting these aside, plaintiffs still enjoy tremendous success under Section 1201. However, when looking at opinions only outside of the Ninth Circuit, win rates become mostly even.

I conclude with several calls for DMCA reform. These include bolstering statutory exceptions and more closely tying Section 1201 to copyright infringement. Pursuing these reforms, I argue, will more faithfully align Section 1201 with its purported objectives.

Keane Woods on Public Law and Private Platforms

Andrew Keane Woods (University of Arizona Law) has posted “Public Law, Private Platforms” (107 Minnesota Law Review 124 (2023)) on SSRN. Here is the abstract:

Our law—both our constitutional law and much of our statutory law—has long drawn a fraught distinction between public and private domains. Indeed, debates about the public/private distinction date as far back as liberalism itself. But today’s private digital platforms strain that distinction to a new degree. Platforms have become our public spaces, but because they are privately owned and “merely” coordinate private ordering, they operate without the guardrails of many of our most important laws.

For example, anti-discrimination law once covered nearly all short-term bookings at inns and hotels; today, nearly a quarter of the hospitality market is controlled by Airbnb, where the majority of bookings are in owner-occupied homes that are exempt from anti-discrimination law’s reach. The First Amendment once protected against the gravest threats to free speech; today, scholars question whether it is fit to handle the novel speech problems presented by social media. The Fourth Amendment once prevented the police from gaining warrantless access to our most private information; today, the police simply buy that data on the open market. The list goes on. While criminal law and speech scholars have noticed the state action problem in constitutional law, and civil rights scholars have discussed the private carveouts in anti-discrimination law, there is little scholarship moving beyond these silos to explore how these different regulatory puzzles stem from the same fundamental problem.

Recognizing that the public/private distinction is the core of the platform problem has a number of implications. It helps explain the platforms’ persistent ability to evade meaningful regulation and it suggests a new way forward—a more suitable remedy than using blunt antitrust tools to address our biggest social ills. Specifically, courts and legislators should revive and expand the legal doctrines that recognize the imperfect nature of our law’s distinction between public and private. These private-but-public doctrines—like public accommodations, the public policy doctrine in contract, the public trust doctrine in property, and more—have long recognized the limits to private ordering in the public interest. It is time to update them for the digital age.

Sokol on Technology Driven Government Law and Regulation

D. Daniel Sokol (USC Gould School of Law) has posted “Technology Driven Government Law and Regulation” (26 Virginia Journal of Law and Technology 1 (2023)) on SSRN. Here is the abstract:

Digitization and digital transformation provides a shock for government to reconceptualize how it is organized to better optimize legal and regulatory responses to the use of data analytics to create value. Government is well situated for an organizational transformation that will allow it to better orchestrate coordinated responses where appropriate based on expertise of a dedicated centralized data analytics unit that will work across agencies. This is not to argue that each government agency should not develop its own data analytics expertise. Rather, there is some expertise that can be leveraged across different parts of government. This type of intervention requires a unique group to coordinate a response.

Katz, Hartung, Gerlach, Jana & Bommarito on NLP in the Legal Domain

Daniel Martin Katz (Illinois Tech – Chicago Kent College of Law; Bucerius Center for Legal Technology & Data Science; Stanford CodeX – The Center for Legal Informatics; 273 Ventures), Dirk Hartung (Bucerius Law School – Center for Legal Technology and Data Science; Stanford University – Stanford Codex Center), Lauritz Gerlach (Bucerius Law School), Abhik Jana
(University of Hamburg; Language Technology Group, Department of Informatics, Universität Hamburg), and Michael James Bommarito (273 Ventures; Licensio, LLC; Stanford Center for Legal Informatics; Michigan State College of Law; Bommarito Consulting, LLC) have posted “Natural Language Processing in the Legal Domain” on SSRN. Here is the abstract:

In this paper, we summarize the current state of the field of NLP and Law with a specific focus on recent technical and substantive developments. To support our analysis, we construct and analyze a corpus of more than six hundred NLP and Law related papers published over the past decade. Our analysis highlights several major trends. Namely, we document an increasing number of papers written, tasks undertaken, and languages covered over the course of the past decade. We observe an increase in the sophistication of the methods which researchers deployed in this applied context. Slowly but surely, Legal NLP is beginning to match the methodological sophistication of general NLP. We believe this to be a positive trend for the future of the field, but many questions in both the academic and commercial sphere still remain open.