I see "This plan may include ads" under the $8 Go tier (accessed https://chatgpt.com/pricing/ from the US) going back to January. Is the behavior change recent?
I think curving has its place. One of my math professors explained that in his opinion an effective test should differentiate performance as much as possible. The top students should score very well and the bottom students should score very poorly. If all the scores are clustered near the top (>80% for example) then it's hard to tell who really mastered the material and who just muddled through. Then, once you've sorted the students you can apply an appropriate curve. He did not have pre-defined thresholds, for each exam he would evaluate when he felt like the quality of work changed from an A to an A-, A- to B+ etc. The curves were very fair; he wasn't trying to force some number of As Bs or Fs, but it did increase my stress levels not knowing in advance how well I needed to do on each exam
Yes, curving has a place here - and it is to evaluate, as you put it, whether the test differentiates performance as much as possible.
If you curve the students after the test, you are applying subjective edits to the graded performance just so the distribution of grades matches the measure of your tests effectiveness. That's just hacking the metric.
Further, even if you believe that tests should differentiate mastery (not students), your test should have teased out the differences or given you enough confidence to provide As to everyone who mastered the material - which should be absolutely possible! There's no a priori reason that all students cannot absolutely get the same grade, except for the a priori assumption that grades are for differentiation of students themselves (this year's A means this is the best student of this year), vs indicating mastery (all students absolutely crushed this exam).
You can dock points for style, or unnecessary struggle, or whatever subjective metric you want, but fudging the grades based on vibes to fit a prior-assumed distribution is just kinda "test effectiveness laundering"
"Best" may include some new posts, I actually haven't checked, but the thing that stands out to me is how old many of the posts are. Ever since Reddit made "best" the default sort on the app I notice that any new subreddit I go to will show me at least some posts from more than two weeks ago. It's really baffling that Reddit seems to think it should be preferred over "hot".
I was going to comment I've never seen this before. Then I went to a subreddit and it hit me that I only use old reddit, which lacks the "best" feature.
I'm not sure why people tolerate the new reddit website. It is so slow, busy, and chock full of ads. When I open it on the phone you are straight up missing a lot of discussion comments, so it is broken too. You click a comment it looks like it has no subcomments under it. You click the permalink opening that comment thread in a new window, and there's still no subcomment under it. Now you prefix old. to the url, now you see the subcomments.
Makes me wonder how much discussion there is that is just not observed at all by a good fraction of the site who browses these same threads. Two universes on the same post.
I prefer the new design because the text is readable on my phone, and it has a native dark mode. I don't see any of the ads because I use an ad blocker.
Also, I have no idea what you mean about the comments thing. I can immediately see all comments other than the ones that have been collapsed due to having negative karma.
I don't know how else to describe it beyond what I've already done. Are you sure you are seeing all the comments? Have you tested with a post? I don't use the app fwiw. Only tested with the mobile website and it has been like this for years.
I still use old.reddit.com on the phone. Works fine with pinch and zoom. Super performant too, loads in a fraction of the time which is necessary with mobile connections and spotty coverage. I just saw the native reddit app for ios at least is like 450mb. WTF...
I don't use the native apps. They are bloated, disgusting messes. I just messed around with old.reddit and compared it to new reddit for around thirty minutes. I noticed no performance differences, and I noticed no discrepancy in comments. I don't know what to tell you. The only time I see fewer comments is when I use an incognito window, but that's just because it's not logged in.
I open this link on my iphone and I only see the top comment by /u/TallGreenhouseGuy. Below that comment, there is a link to "more replies" but it loads this exact same page with only the parent comment, no child comments. Below the parent comment, is just random reddit threads, absolutely random, along with ads. Top 3: post from /r/ghanacitizen (I am not in africa no clue how that is there....), ad from cerave for oil control shampoo, then a link to some thread in /r/learnprogramming.
I agree, and I previously used a 3rd party app myself (Reddit Is Fun). The sad thing is that I actually completely understand why they shut down the free API access. The AI bot scrapers are absurdly aggressive, and bandwidth isn't free.
For count 3, the prediction markets consider the "bets" to actually be futures contracts, and futures contracts are regulated together with commodities (in the U.S. by the CFTC). There is ongoing litigation about whether this is the proper designation, but that is the U.S. government's position. Insider trading rules are more lax for futures than other products, but I believe this case likely does violate existing rules.
Anyone who treats Geekbench as a meaningful benchmark (i.e. not without a huge disclaimer or with other more meaningful datapoints) is not to be trusted. You can only really trust it for inter-generational comparisons within a single architecture.
Not true. Geekbench, especially single threaded benchmark, is probably the best we got, it has a bunch of workloads, unlike many other benchmarks like cinebench for example. And they publish all the results on their website, so you can dig into each individual workload and find the ones that apply to you.
And like the other poster mentioned, it correlates well with SPEC, so it's basically a easily accessible SPEC. These days the only benchmark I use to quickly judge some CPU is geekbench.
May I suggest the one I use (I wrote it), which also correlates well with SPEC & Geekbench 5, but also runs the benchmarks on all cores if you want to so you get both max single-thread and max multi-thread: https://github.com/dkechag/dkbench-docker . You basically run 'docker run -it --rm dkechag/dkbench'.
I took a look, it's not bad but it seems to contain too many micro benchmarks like regex or primes. Geekbench at least has clang which is a subscore that I always look at.
The primes one is my least favourite one indeed, I left it in just because I happened to include it in the very first version and I am thinking it just counts for 5% in the end...
The regex ones are "micro" yet quite important, dkbench it's a Perl (and C)-based benchmark (reflects our main code), and the regex engine is the most highly optimized part of the language so regex speed is a good representation of text processing speed in Perl.
As I said, the overall score correlates well to SPEC/Geekbench so as a suite it works well.
For compiler comparisons I usually compile a language like python or perl as a test, but I did not want to add something like that, to keep it fast with many smaller benchmarks.
I think it's that assumption is the problem. Most social systems are predicated on having enough net contributors to provide for net recipients, but with a declining population the ratio of contributors/recipients can get small. There may be solutions to this, but current social systems will likely fail if left unchanged. That doesn't mean the only solution is population growth, but we do need to do something
I can't say for sure about the Wang terminal keyboards, but what you're describing sounds a lot like a mechanism from some IBM Model B keyboards (usually called Beamsprings). I have an IBM 5251 keyboard that has a solenoid that hammers the side of the metal case whenever you type, and I've heard that it was added as users would have been used to typewriters and wanted to know for sure when they had registered a keypress
So honestly I don't quite remember if I encountered this with the Wangs, or if I'm recalling my Dad telling me about it from his experiences.
If the latter then odds are that it was either a machine from Wang and in that case most likely the 2200, or otherwise it will have most probably been equipment associated with the Gamma 10 from De La Rue Bull, or possibly the Ferranti Pegasus - both of which I know he worked with.
Of course, he might have been telling me a third-party anecdote in which case it's possible the IBM Display Station was the machine in question.
That all said, last time I was discussing this with someone they mentioned that the 2200's terminal had a "solenoid" trace on its PCB so it's quite possible that this really was the relevant device. Last time I personally had hands on a live 2200 was about 1993 though, so I really can't be sure.
There's a chap in the Netherlands with a Wang 2200 museum - perhaps I should just write to him and ask :D
I haven't looked at any court documents, but the WSJ article from Wednesday reported that "Last year, Google sued the anonymous operators of a network of more than 10 million internet-connected televisions, tablets and projectors, saying they had secretly pre-installed residential proxy software on them... an Ipidea spokeswoman acknowledged in an email that the company and its partners had engaged in “relatively aggressive market expansion strategies” and “conducted promotional activities in inappropriate venues (e.g., hacker forums)...”"
There was also a botnet, Kimwolf, that apparently leveraged an exploit to use the residential proxy service, so it may be related to Ipidea not shutting them down.
VK_OEM_MINUS 0xBD For any country/region, the Dash and Underscore key
VK_SUBTRACT 0x6D Subtract key
I'm not sure if I've ever heard someone say "Subtract key" but I guess it makes sense for it's purpose on the numpad
This other source has a different description: https://learn.microsoft.com/en-us/dotnet/api/system.windows....
OemMinus 189 The OEM minus key on any country/region keyboard.